Model Details Model Card for Olmo 3.1 32B Instruct We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain of thought thinking improves reasoning tasks like math and coding. Olmo is a series of O pen l anguage mo dels designed to enable the science of language models. These models are pre trained on the Dolma 3 dataset and post trained on the Dolci datasets. We are releasing all code, checkpoints, logs (coming soon), and associated training details. The core models released in this batch include the following: Stage Olmo 3 7B Think Olmo (3/3.1) 32B Think Olmo 3 7B Instruct Olmo 3.1 32B Instruct Base Model Olmo 3 7B Olmo 3 32B Olmo 3 7B Olmo 3 32B SFT Olmo 3 7B Think SFT Olmo 3 32B Think SFT Olmo 3 7B Instruct SFT Olmo 3.1 32B Instruct SFT DPO Olmo 3 7B Think DPO Olmo 3 32B Think DPO Olmo 3 7B Instruct DPO Olmo 3.1 32B Instruct DPO Final Models (RLVR) Olmo 3 7B Think Olmo 3 32B Think Olmo 3.1 32B Think Olmo 3 7B Instruct Olmo 3.1 32B Instruct Installation Olmo 3 is supported in transformers 4.57.0 or higher: Inference You can use OLMo with the standard HuggingFace transformers library: For faster performance, you can quantize the model usi…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy