Moonshot AI’s Kimi K3 (Max) just did something that open-weight models haven’t managed before: it climbed into the top three of Agent Arena’s overall rankings, sitting behind only Claude Fable 5 (High) and GPT-5.6 Sol (xHigh). With a +9.75% net-improvement score, it’s the highest-performing open-weight model across all 42 evaluated systems.
What makes K3 different
Kimi K3 is the first open-weight model in the multi-trillion-parameter class, packing 2.8 trillion parameters in a Mixture-of-Experts architecture. Think of MoE like a hospital with specialist doctors: instead of routing every patient through a general practitioner, the system activates only the relevant experts for each task. This keeps compute costs manageable despite the massive parameter count.
The model ships with a 1 million token context window and native vision capabilities. Moonshot AI claims K3 achieves 2.5 times the intelligence-per-compute efficiency compared to its predecessor, K2.
The model went live publicly around July 14-16, with full weights and a technical report following on July 27.











