Ouro 2.6B Thinking Model Description ⚠️ IMPORTANT: This model is intended for research purposes only. It is provided as is without warranties for production use. Ouro 2.6B Thinking is a reasoning specialized variant of the Ouro 2.6B base model, enhanced through supervised fine tuning on high quality reasoning data. Please use transformers==4.54.1 for compatibility. Key Features Advanced Reasoning : Specifically optimized for mathematical and scientific reasoning tasks Compact Size : Competitive with 4B models despite having only 2.6B parameters Cross Step Consistency : Intermediate recurrent outputs can serve as reliable proxies for final answers Explicit Thinking Process : Trained to generate detailed reasoning steps Configuration Recurrent Steps and Adaptive Exit The model's computational behavior can be configured through the config.json file: total ut steps : Controls the number of recurrent steps (default: 4). You can adjust this value to trade off between performance and computation time. early exit threshold : Controls the adaptive exit mechanism (default: 1.0). Lower values encourage earlier exit, while 1.0 means always use all steps. Example: Modify recurrent steps Note :…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy