Kimi K3 ranks second only to Fable 5 on agentic benchmark

Moonshot's 2.8T-parameter Kimi K3 scores 57 on Artificial Analysis and ranks second to Fable 5 on AA-Briefcase; it costs more than Opus 4.8 to run and averages nearly an hour per task.
How this story unfolded
5 days · 0 reports · 4 community posts · 4 of 5 shown
Moonshot AI by email
Get an email when Moonshot AI has news
No news that day, no email.
More stories today
- Anton: self-improving terminal AI agent automates inbox, calendar, reports
- Allie Mellen discusses AI's cybersecurity impact at Black Hat 2026
- Satirical post by Timnit Gebru mocks 'autonomous AGI startup' hype
- AI YouTube Shorts Generator turns long videos into vertical Shorts
- Domain name tool generates 60 creative startup name candidates