[!NOTE] Includes our GGUF chat template fixes ! Tool calling works as well! If you are using llama.cpp , use jinja to enable the system prompt. Unsloth Dynamic 2.0 achieves SOTA performance in model quantization. ✨ How to Use Mistral 3.2 Small: Run in llama.cpp: Run in Ollama: Temperature of: 0.15 Set top p to: 1.00 Max tokens (context length): 128K Fine tune Mistral v0.3 (7B) for free using our Google Colab notebook here Conversational.ipynb)! View the rest of our notebooks in our docs here. Mistral Small 3.2 24B Instruct 2506 Mistral Small 3.2 24B Instruct 2506 is a minor update of Mistral Small 3.1 24B Instruct 2503. Small 3.2 improves in the following categories: Instruction following : Small 3.2 is better at following precise instructions Repetition errors : Small 3.2 produces less infinite generations or repetitive answers Function calling : Small 3.2's function calling template is more robust (see here and examples) In all other categories Small 3.2 should match or slightly improve compared to Mistral Small 3.1 24B Instruct 2503. Key Features same as Mistral Small 3.1 24B Instruct 2503 Benchmark Results We compare Mistral Small 3.2 24B to Mistral Small 3.1 24B Instruct 2503.…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy