Moonshot's open-weight Kimi K3 beats Opus 4.8 on frontier benchmarks

Moonshot launched Kimi K3 on July 17 with 2.8 trillion parameters, fully open-source; Artificial Analysis ranks it ahead of Anthropic's Opus 4.8 on frontier benchmarks — the first Chinese open-weight model to do so. It still trails Claude Fable 5 and GPT-5.6 overall.
15 sources
Kimi K3: The open-weights escalationinterconnects.ai
Kimi K3 is second only to Fable 5 on AA-Briefcase, our agentic knowledge work benchmark, but costs...x.com
A lot of swift conclusions are being drawn about Kimi K3 based on fairly saturated benchmarks and EL...bsky.app
Kimi K3 vs DeepSeek V4 Pro vs GLM-5.2: Open Trillion-Scale MoE Models Compared on Benchmarks, License, and Serving Costmarktechpost.com
Kimi K3 tops Arena’s coding leaderboard — and it’s open-weightthenewstack.io
KIMI K3 Beats Claude Fable and GPT 5.6 sol in arena.ai!!!reddit.com
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Cogent AI releases VR-1 cyber reasoning model
- Orchestrator tool integrates 12 AI coding agents in Visual Studio Code
- Taste Skill rules file for AI coding agents crosses 70,000 GitHub stars
- Hugging Face Diffusers flaws allow arbitrary code execution
- 139 Agent Skills bring legal workflows to Claude, Codex and Gemini CLI