Qwen 3.8-Max and Claude Opus 5: benchmark scores don't predict the bill

Alibaba released Qwen 3.8-Max this week, marketing the preview as second only to Claude Fable 5, though the launch-day table had it leading on just one of 12 coding-agent rows. An independent benchmark harness reportedly came close to the opposite conclusion.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Sequoia Capital invests in AI-native video platform Preview
- US Launches Effort to Speed Trade in AI Goods Between Allies
- DeepMind launches SL2T sign language-to-text model
- Liquid AI releases LFM2.5-VL-3B vision-language model for edge
- Grok and Meta's release discussed on ETN podcast episode