Z.ai releases GLM-5.3 Flash, a 320B-parameter multimodal model

GLM-5.3 Flash packs 320B parameters (18B active), 1M context, and hybrid attention. On Databricks' OfficeQA Pro v2, it delivers 10% higher quality than GLM-5.2 at one-tenth the cost, and nearly matches Luna's DeepSWE performance at twice the throughput.
How this story unfolded
3 days · 3 reports · 5 community posts · 8 of 9 shown
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- DeepMind panel discusses generative media SOTA and human eval
- Reddit users share useful MCP servers for Claude
- Why embodied AI hits an edge AI wall requiring new math
- HTMX CEO mandates 'No AI Fridays' to counter LLM cognitive debt
- OpenAI product lead Tara Seshan discusses persistent AI coworkers