paper dataset model [!NOTE] We have released a paper for OpenThoughts! See our paper here. OpenThinker3 7B State of the art open data 7B reasoning model. ๐ This model is a fine tuned version of Qwen/Qwen2.5 7B Instruct on the OpenThoughts3 1.2M dataset. It represents a notable improvement over our previous models, OpenThinker 7B and OpenThinker2 7B, and it outperforms several other strong reasoning 7B models such as DeepSeek R1 Distill Qwen 7B and Llama 3.1 Nemotron Nano 8B v1, despite being trained only with SFT, without any RL. This time, we also released a paper! See our paper and blog post for more details. OpenThinker3 32B to follow! ๐ Evaluation Results The numbers reported in the table below are evaluated with our open source tool Evalchemy. In the table below, we bold values in each column that are within 2 standard errors of the best. Model Data AIME24 AIME25 AMC23 MATH500 HMMT O2/25 LCB 06/24 01/25 CodeElo CodeForces GPQA D JEEBench OpenThinker 7B โ 30.7 22.0 72.5 82.8 15.7 26.1 11.1 14.9 38.6 45.3 OpenThinker2 7B โ 60.7 38.7 89.8 87.6 24.7 40.6 22.8 26.6 47.0 65.1 OpenThinker3 7B โ 69.0 53.3 93.5 90.0 42.7 51.7 31.0 32.2 53.7 72.4 DeepSeek R1 Distill Qwen 32B โ 51.3 38โฆ
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy