[!NOTE] Includes Unsloth chat template fixes ! For llama.cpp , use jinja Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants. Cogito v2 preview 109B MoE Blog Post The Cogito v2 LLMs are instruction tuned generative models. All models are released under an open license for commercial use. Cogito v2 models are hybrid reasoning models. Each model can answer directly (standard LLM), or self reflect before answering (like reasoning models). The LLMs are trained using Iterated Distillation and Amplification (IDA) an scalable and efficient alignment strategy for superintelligence using iterative self improvement. The models have been optimized for coding, STEM, instruction following and general helpfulness, and have significantly higher multilingual, coding and tool calling capabilities than size equivalent counterparts. In both standard and reasoning modes, Cogito v2 preview models outperform their size equivalent counterparts on common industry benchmarks. This model is trained in over 30 languages and supports long contexts (upto 10M tokens). Evaluations For detailed evaluations, please refer to the Blog Post. Usage Here is a snippet below for usage with T…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy