OLMo 2 32B Instruct March 2025 is post trained variant of the OLMo 2 32B March 2025 model, which has undergone supervised finetuning on an OLMo specific variant of the Tülu 3 dataset, further DPO training on this dataset, and final RLVR training on this dataset. Tülu 3 is designed for state of the art performance on a diversity of tasks in addition to chat, such as MATH, GSM8K, and IFEval. Check out the OLMo 2 paper or Tülu 3 paper for more details! OLMo is a series of O pen L anguage Mo dels designed to enable the science of language models. These models are trained on the Dolma dataset. We are releasing all code, checkpoints, logs, and associated training details. Model description Model type: A model trained on a mix of publicly available, synthetic and human created datasets. Language(s) (NLP): Primarily English License: Apache 2.0 Finetuned from model: allenai/OLMo 2 0325 32B DPO Model Sources Project Page: https://allenai.org/olmo Repositories: Core repo (training, inference, fine tuning etc.): https://github.com/allenai/OLMo core Evaluation code: https://github.com/allenai/olmes Further fine tuning code: https://github.com/allenai/open instruct Paper: https://arxiv.org/abs/2…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy