Ouro 1.4B Thinking Model Description ⚠️ IMPORTANT: This model is intended for research purposes only. It is provided as is without warranties for production use. Ouro 1.4B Thinking is a reasoning specialized variant of the Ouro 1.4B base model, enhanced through supervised fine tuning on high quality reasoning data. Key Features Advanced Reasoning : Specifically optimized for mathematical and scientific reasoning tasks Compact Size : Competitive with 4B models despite having only 1.4B parameters Cross Step Consistency : Intermediate recurrent outputs can serve as reliable proxies for final answers Explicit Thinking Process : Trained to generate detailed reasoning steps Configuration Recurrent Steps and Adaptive Exit The model's computational behavior can be configured through the config.json file: total ut steps : Controls the number of recurrent steps (default: 4). You can adjust this value to trade off between performance and computation time. early exit threshold : Controls the adaptive exit mechanism (default: 1.0). Lower values encourage earlier exit, while 1.0 means always use all steps. Example: Modify recurrent steps Note : vLLM does not currently support the adaptive exit f…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy