← Back to brief
ModelsReportedThe Decoder

Moonshot's Kimi K3 tops frontend code rankings but trails in advanced math

Moonshot's Kimi K3 is the first Chinese model to lead the Code Arena: Frontend rankings, outperforming Claude Fable 5 and GPT-5.6 Sol. However, on FrontierMath Tier 4, Kimi K3 scores only about 39%, while OpenAI and Anthropic models achieve close to 90%.

Why it matters: This highlights a significant gap in advanced mathematical reasoning between leading Chinese and Western AI models, even as Chinese models excel in specific coding tasks.

Full story at: The Decoder