AnalysisAI ModelsAugust 29, 2026

GLM-5.3 Flash: 17x cheaper, 5.6 points behind on DeepSWE

Together AI ran 900 DeepSWE rollouts: GLM-5.3 Flash trails the full model by 5.6 points at pass@1 but only 2.6 at pass@4, at 17x lower cost ($0.24 vs $3.99 per rollout). A cascade routing solves 80.9% of tasks at $1.70 each.

How this story unfolded

3 days · 4 reports · 6 community posts · from Aug 29

  1. Aug 29
  2. Aug 30
  3. Sep 1

AI Models by email

Get an email when there's news on AI Models

No news that day, no email.

More stories today

Open the live feed