π Qwen3.5 9B Gemini 3.1 Pro Reasoning Distill π‘ Model Introduction Qwen3.5 9B Gemini 3.1 Pro Reasoning Distill is a reasoning model fine tuned on top of Qwen3.5 9B . The model is primarily optimized through high density reasoning distillation sourced from Gemini 3.1 , while also incorporating additional reasoning traces distilled from Qwen3.5 27B and a broader Gemini 3.0 Pro reasoning corpus. Through Supervised Fine Tuning focused on structured analytical behavior, this model aims to reshape the base modelβs reasoning style into a more coherent, better organized, and higher density Chain of Thought (CoT) pattern. It is especially designed to improve decomposition, planning, abstraction, and response cleanliness on complex multi step tasks. π§ Example of Learned Reasoning Scaffold This model inherits a more structured reasoning style influenced by Gemini 3.1 style analytical planning . Compared with more loosely exploratory reasoning patterns, this model tends to organize the problem before answering: πΊοΈ Training Pipeline Overview π Stage Details πΉ Supervised Fine Tuning (SFT) Objective: Objective: To inject reasoning behavior into Qwen3.5 9B and strengthen its performance on cβ¦
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy