Model Overview Description: Llama 3.1 Nemotron 70B Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries. This model reaches Arena Hard of 85.0, AlpacaEval 2 LC of 57.6 and GPT 4 Turbo MT Bench of 8.98, which are known to be predictive of LMSys Chatbot Arena Elo As of 1 Oct 2024, this model is 1 on all three automatic alignment benchmarks (verified tab for AlpacaEval 2 LC), edging out strong frontier models such as GPT 4o and Claude 3.5 Sonnet. As of Oct 24th, 2024 the model has Elo Score of 1267(+ 7), rank 9 and style controlled rank of 26 on ChatBot Arena leaderboard. This model was trained using RLHF (specifically, REINFORCE), Llama 3.1 Nemotron 70B Reward and HelpSteer2 Preference prompts on a Llama 3.1 70B Instruct model as the initial policy. Llama 3.1 Nemotron 70B Instruct HF has been converted from Llama 3.1 Nemotron 70B Instruct to support it in the HuggingFace Transformers codebase. Please note that evaluation results might be slightly different from the Llama 3.1 Nemotron 70B Instruct as evaluated in NeMo Aligner, which the evaluation results below are based on. Try hosted inference for free at build…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy