📰 Tech Blog 📄 Paper 0. Changelog 2025.8.11 Messages with name field are now supported. We’ve also moved the chat template to a standalone file for easier viewing. 2025.7.18 We further modified our chat template to improve its robustness. The default system prompt has also been updated. 2025.7.15 We have updated our tokenizer implementation. Now special tokens like [EOS] can be encoded to their token ids. We fixed a bug in the chat template that was breaking multi turn tool calls. 1. Model Introduction Kimi K2 is a state of the art mixture of experts (MoE) language model with 32 billion activated parameters and 1 trillion total parameters. Trained with the Muon optimizer, Kimi K2 achieves exceptional performance across frontier knowledge, reasoning, and coding tasks while being meticulously optimized for agentic capabilities. Key Features Large Scale Training: Pre trained a 1T parameter MoE model on 15.5T tokens with zero training instability. MuonClip Optimizer: We apply the Muon optimizer to an unprecedented scale, and develop novel optimization techniques to resolve instabilities while scaling up. Agentic Intelligenc…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy