GLM-5.3 matches frontier models on DeepSWE at fraction of cost

Together AI ran 904 DeepSWE rollouts: GLM-5.3 ties Claude Fable 5 on pass@1 (69.0% vs 69.7%) and beats GPT-5.6 Sol at pass@4 (87.6% vs 85.8%), at $3.99 per rollout vs $21.63 and $8.37. A GLM-first cascade hits 85.9% at $6.61 per task.
How this story unfolded
2 days · 2 reports · 4 community posts · from Aug 22
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Anthropic co-founder: chips, not algorithms, bottleneck AI
- Teachers targeted by sexualized AI deepfakes from students
- FDA promises generative AI medical device guidance
- ConvRot quantization method lands in llama-cpp-turboquant
- Hermes adds auxiliary review model to /review command