Nous Hermes 2 Mixtral 8x7B DPO Model description Nous Hermes 2 Mixtral 8x7B DPO is the new flagship Nous Research model trained over the Mixtral 8x7B MoE LLM. The model was trained on over 1,000,000 entries of primarily GPT 4 generated data, as well as other high quality data from open datasets across the AI landscape, achieving state of the art performance on a variety of tasks. This is the SFT + DPO version of Mixtral Hermes 2, we have also released an SFT only version, for people to find which works best for them, which can be found here: https://huggingface.co/NousResearch/Nous Hermes 2 Mixtral 8x7B SFT We are grateful to Together.ai for sponsoring our compute during the many experiments both training Mixtral and working on DPO! Table of Contents 1. Example Outputs 2. Benchmark Results GPT4All AGIEval BigBench Comparison to Mixtral Instruct 3. Prompt Format 4. Inference Example Code 5. Quantized Models Example Outputs Writing Code for Data Visualization Writing Cyberpunk Psychedelic Poems Performing Backtranslation to Create Prompts from Input Text Benchmark Results Nous Hermes 2 on Mixtral 8x7B is a major improvement across the board on the benchmarks below compared to the bas…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy