Cat llama3 instruct Abstract We present cat llama3 instruct, a llama 3 70b finetuned model focusing on system prompt fidelity, helpfulness and character engagement. The model aims to respect system prompt to an extreme degree, and provide helpful information regardless of situations and offer maximum character immersion(Role Play) in given scenes. Introduction Llama 3 70b provides a brand new platform that’s more knowledgeable and steerable than the previous generations of products. However, there currently lacks general purpose finetunes for the 70b version model. Cat llama3 instruct 70b aims to address the shortcomings of traditional models by applying heavy filtrations for helpfulness, summarization for system/character card fidelity, and paraphrase for character immersion. Specific Aims: System Instruction fidelity Chain of Thought(COT) Character immersion Helpfulness for biosciences and general science Methods Dataset Preparation Huggingface dataset containing instruction response pairs was systematically pulled. We have trained a gpt model on gpt4 responses exclusively to serve as a standard model. (Fig1. Huggingface dataset population distribution and filtration for each com…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy