Infinity Instruct Beijing Academy of Artificial Intelligence (BAAI) [Paper][Code][π€] (would be released soon) Infinity Instruct 3M 0625 Llama3 8B is an opensource supervised instruction tuning model without reinforcement learning from human feedback (RLHF). This model is just finetuned on Infinity Instruct 3M and Infinity Instruct 0625 and showing favorable results on AlpacaEval 2.0 and MT Bench. News π₯π₯π₯[2024/07/09] We release the model weights of InfInstruct Mistral 7B 0625, InfInstruct Qwen2 7B 0625, InfInstruct Llama3 8B 0625, InfInstruct Llama3 70B 0625, and InfInstruct Yi 1.5 9B 0625. π₯π₯π₯[2024/07/09] We release the chat dataset Infinity Instruct 0625, it is a upgraded version of the Infinity Instruct 0613. π₯π₯π₯[2024/06/28] We release the model weight of InfInstruct Llama3 70B 0613. It shows favorable results on AlpacaEval 2.0 compared to GPT4 0613 without RLHF. π₯π₯π₯[2024/06/21] We release the model weight of InfInstruct Mistral 7B 0613. It shows favorable results on AlpacaEval 2.0 compared to Mixtral 8x7B v0.1, Gemini Pro, and GPT 3.5 without RLHF. π₯π₯π₯[2024/06/13] We share the intermediate result of our data construction process (corresponding to the InfInstructβ¦
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy