Llamacpp imatrix Quantizations of Mistral Small 24B Instruct 2501 Using llama.cpp release b4585 for quantization. Original model: https://huggingface.co/mistralai/Mistral Small 24B Instruct 2501 All quants made using imatrix option with dataset from here Run them in LM Studio Run them directly with llama.cpp, or any other llama.cpp based project Prompt format Download a file (not the whole branch) from below: Filename Quant type File Size Split Description Mistral Small 24B Instruct 2501 f32.gguf f32 94.30GB true Full F32 weights. Mistral Small 24B Instruct 2501 f16.gguf f16 47.15GB false Full F16 weights. Mistral Small 24B Instruct 2501 Q8 0.gguf Q8 0 25.05GB false Extremely high quality, generally unneeded but max available quant. Mistral Small 24B Instruct 2501 Q6 K L.gguf Q6 K L 19.67GB false Uses Q8 0 for embed and output weights. Very high quality, near perfect, recommended . Mistral Small 24B Instruct 2501 Q6 K.gguf Q6 K 19.35GB false Very high quality, near perfect, recommended . Mistral Small 24B Instruct 2501 Q5 K L.gguf Q5 K L 17.18GB false Uses Q8 0 for embed and output weights. High quality, recommended . Mistral Small 24B Instruct 2501 Q5 K M.gguf Q5 K M 16.76GB false…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy