Moonshot's Kimi K3 tops frontend code rankings but trails in advanced math
Moonshot's Kimi K3 is the first Chinese model to lead the Code Arena: Frontend rankings, outperforming Claude Fable 5 and GPT-5.6 Sol. However, on FrontierMath Tier 4, Kimi K3 scores only about 39%, while OpenAI and Anthropic models achieve close to 90%.
Why it matters: This highlights a significant gap in advanced mathematical reasoning between leading Chinese and Western AI models, even as Chinese models excel in specific coding tasks.
Full story at: The Decoder ↗