Mistral Small 4 119B A6B Mistral Small 4 is a powerful hybrid model capable of acting as both a general instruction model and a reasoning model. It unifies the capabilities of three different model families— Instruct , Reasoning (previously called Magistral), and Devstral —into a single, unified model. With its multimodal capabilities, efficient architecture, and flexible mode switching, it is a powerful general purpose model for any task. In a latency optimized setup, Mistral Small 4 achieves a 40% reduction in end to end completion time , and in a throughput optimized setup, it handles 3x more requests per second compared to Mistral Small 3. To further improve efficiency you can either take advantages of: Speculative decoding thanks to our trained eagle head mistralai/Mistral Small 4 119B 2603 eagle . 4 bit float precision quantization thanks to our NVFP4 checkpoint mistralai/Mistral Small 4 119B 2603 NVFP4 . Key Features Mistral Small 4 includes the following architectural choices: MoE : 128 experts, 4 active. 119B parameters , with 6.5B activated per token . 256k context length . Multimodal input : Accepts both text and image input, with text output. Instruct and Reasoning func…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy