OLMo 2 1124 13B Instruct NOTE: 1/3/2025 UPDATE: Upon the initial release of OLMo 2 models, we realized the post trained models did not share the pre tokenization logic that the base models use. As a result, we have trained new post trained models. The new models are available under the same names as the original models, but we have made the old models available with a postfix " preview". See OLMo 2 Preview Post trained Models for the colleciton of the legacy models. Release Documentation OLMo 2 13B Instruct November 2024 is post trained variant of the OLMo 2 13B November 2024 model, which has undergone supervised finetuning on an OLMo specific variant of the Tülu 3 dataset, and finally RLVR training using this data. Tülu 3 is designed for state of the art performance on a diversity of tasks in addition to chat, such as MATH, GSM8K, and IFEval. Check out the OLMo 2 paper or Tülu 3 paper for more details! OLMo is a series of O pen L anguage Mo dels designed to enable the science of language models. These models are trained on the Dolma dataset. We are releasing all code, checkpoints, logs (coming soon), and associated training details. The core models released in this batch include t…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy