Qwen3Guard Gen 8B Qwen3Guard is a series of safety moderation models built upon Qwen3 and trained on a dataset of 1.19 million prompts and responses labeled for safety. The series includes models of three sizes (0.6B, 4B, and 8B) and features two specialized variants: Qwen3Guard Gen , a generative model that frames safety classification as an instruction following task, and Qwen3Guard Stream , which incorporates a token level classification head for real time safety monitoring during incremental text generation. This repository hosts Qwen3Guard Gen , which offers the following key advantages: Three Tiered Severity Classification: Enables detailed risk assessment by categorizing outputs into safe, controversial, and unsafe severity levels, supporting adaptation to diverse deployment scenarios. Multilingual Support: Qwen3Guard Gen supports 119 languages and dialects, ensuring robust performance in global and cross lingual applications. Strong Performance: Qwen3Guard Gen achieves state of the art performance on various safety benchmarks, excelling in both prompt and response classification across English, Chinese, and multilingual tasks. For more details, please refer to our blog, Git…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy