AnalysisAI ModelsJuly 9, 2026
Grok 4.5 rivals Claude Opus 4.8 on agentic coding benchmarks

Grok 4.5, trained on Cursor data, now rivals Claude Opus 4.8 on coding benchmarks. The head-to-head weighs cost, speed, and real-world agentic coding performance.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Record chip earnings fail to satisfy AI investors as stocks slide
- PolyAI releases Dialog-RSN-1 audio-native dialog model
- Build a policy-governed multi-agent financial workflow with Omnigent
- H3 enters arena; MiniMax open weights coming soon
- Thinking Machines releases Inkling-Small multimodal model