[!Note] This is an improved version of Kimi VL A3B Thinking. Please consider using this updated model instead of the previous version. [!Note] Please visit our tech blog for recommended inference recipe of this model: Kimi VL A3B Thinking 2506: A Quick Navigation š Tech Report š Github š¬ Chat Web 1. Introduction This is an updated version of Kimi VL A3B Thinking, with following improved abilities: It Thinks Smarter while Consuming Less Tokens : The 2506 version reaches better accuracy on multimodal reasoning benchmarks: 56.9 on MathVision (+20.1), 80.1 on MathVista (+8.4), 46.3 on MMMU Pro (+3.3), 64.0 on MMMU (+2.1), while in average requires 20\% reduced thinking length. It Sees Clearer with Thinking : Unlike the previous version that specializes on thinking tasks, the 2506 version can also achieve the same or even better ability on general visual perception and understanding, e.g. MMBench EN v1.1 (84.4), MMStar (70.4), RealWorldQA (70.0), MMVet (78.4), surpassing or matching abilties of our non thinking model (Kimi VL A3B Instruct). It Extends to Video Scenarios : The new 2506 version also improves on video reasoning and understanding benchmarks. Iā¦
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy