Infinity Instruct Beijing Academy of Artificial Intelligence (BAAI) [Paper][Code][๐ค] (would be released soon) Infinity Instruct 7M Gen Llama3.1 8B is an opensource supervised instruction tuning model without reinforcement learning from human feedback (RLHF). This model is just finetuned on Infinity Instruct 7M and Infinity Instruct Gen and showing favorable results on AlpacaEval 2.0 compared to GPT4. News ๐ฅ๐ฅ๐ฅ[2024/08/02] We release the model weights of InfInstruct Llama3.1 70B Gen, InfInstruct Llama3.1 8B Gen, InfInstruct Mistral 7B Gen. ๐ฅ๐ฅ๐ฅ[2024/08/02] We release the 7M foundational dataset Infinity Instruct 7M. ๐ฅ๐ฅ๐ฅ[2024/07/09] We release the model weights of InfInstruct Mistral 7B 0625, InfInstruct Qwen2 7B 0625, InfInstruct Llama3 8B 0625, InfInstruct Llama3 8B 0625, and InfInstruct Yi 1.5 9B 0625. ๐ฅ๐ฅ๐ฅ[2024/07/09] We release the chat dataset Infinity Instruct 0625, it is a upgraded version of the Infinity Instruct 0613. ๐ฅ๐ฅ๐ฅ[2024/06/28] We release the model weight of InfInstruct Llama3 8B 0613. It shows favorable results on AlpacaEval 2.0 compared to GPT4 0613 without RLHF. ๐ฅ๐ฅ๐ฅ[2024/06/21] We release the model weight of InfInstruct Mistral 7B 0613. It shows favโฆ
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy