Qwen3 Swallow Qwen3 Swallow v0.2 is a family of large language models available in 8B , 30B A3B , and 32B parameter sizes. Built as bilingual Japanese English models, they were developed through Continual Pre Training (CPT), Supervised Fine Tuning (SFT), and Reinforcement Learning with Verifiable Rewards (RLVR) based on the Qwen3 [Yang, 2025]. In addition to enhancing Japanese language proficiency and Japanese English translation capabilities, we significantly improved or maintained performance on math and coding tasks during CPT by using high quality math and code datasets with reasoning traces, along with custom built datasets during SFT. Subsequently, we further improved the models' math and coding performance by enhancing reasoning capabilities through RLVR. Qwen3 Swallow Project Page [!NOTE] Please note that Qwen3 Swallow v0.1 is a skipped version number. Highlights Bilingual Proficiency: Highly optimized for both Japanese and English. Retained STEM Performance: Strategic CPT and SFT pipelines successfully prevented catastrophic forgetting in mathematics and coding. Enhanced Reasoning: Achieved reasoning performance on par with the original Qwen3 models, and even surpassing th…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy