🔮 Beyonder 4x7B v3 Beyonder 4x7B v3 is an improvement over the popular Beyonder 4x7B v2. It's a Mixture of Experts (MoE) made with the following models using LazyMergekit: mlabonne/AlphaMonarch 7B beowolx/CodeNinja 1.0 OpenChat 7B SanjiWatsuki/Kunoichi DPO v2 7B mlabonne/NeuralDaredevil 7B Special thanks to beowolx for making the best Mistral based code model and to SanjiWatsuki for creating one of the very best RP models. Try the demo : https://huggingface.co/spaces/mlabonne/Beyonder 4x7B v3 🔍 Applications This model uses a context window of 8k. I recommend using it with the Mistral Instruct chat template (works perfectly with LM Studio). If you use SillyTavern, you might want to tweak the inference parameters. Here's what LM Studio uses as a reference: temp 0.8, top k 40, top p 0.95, min p 0.05, repeat penalty 1.1. Thanks to its four experts, it's a well rounded model, capable of achieving most tasks. As two experts are always used to generate an answer, every task benefits from other capabilities, like chat with RP, or math with code. ⚡ Quantized models Thanks bartowski for quantizing this model. GGUF : https://huggingface.co/mlabonne/Beyonder 4x7B v3 GGUF More GGUF : https://…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy