Z.ai releases open-weight GLM-5.3

Z.ai released GLM-5.3 as open-weight, available on Hugging Face and in Perplexity Computer. It scores 84.5% on CyberGym and 60 on the Artificial Analysis Intelligence Index, beating GPT-5.6 and Claude Fable 5 on agentic benchmarks.
How this story unfolded
3 weeks · 17 reports · 34 community posts · 51 of 55 shown
- Aug 14
Z.ai Ships GLM-5.3 Without Retraining the Base Model: Better at Complex Coding and Long-Horizon Tasksmarktechpost.com
GLM-5.3 didn’t change the base model — where did its coding gains come from?thenewstack.io
China's Z.AI Ships GLM-5.3, Calling It the Top Open-Weight Coding Modeldecrypt.co
GLM-5.3: How Chinese labs keep stride with the frontierinterconnects.ai
GLM-5.3 is here with advanced cyber capabilities — and reportedly already found a 'serious vulnerability' in Cursorventurebeat.com
- Aug 15
- Aug 16
- Aug 17
- Aug 18
- Aug 19
GLM-5.3 hits the API at $1.4/$4.4 per million tokensventurebeat.com
An industrial-scale distillation of models, or subtle benchmaxxing: What developers really think of GLM-5.3thenewstack.io
GLM-5.3 achieves 60 on the Artificial Analysis Intelligence Index, on par with Kimi K3 and up 7 points from GLM-5.2. Once the weights are released it will be tied as the leading open weights model
- Aug 28
- Aug 29
- Aug 30
- Aug 31
- Sep 1
- Sep 2
More stories today
Reddit compares Sol 5.6 Ultra and Astra Light outputs
A Reddit user shared side-by-side outputs from Sol 5.6 Ultra and Astra 6 Light in work mode, calling both "insane" and saying they wouldn't need anything above Astra Light for real work.
r/Singularity·1 hour ago
Anthropic resets Claude Code usage limits
Anthropic has reset usage limits for Claude Code, restoring full access for users. The reset was confirmed by multiple users on social media on September 4, 2026.
TestingCatalog News·1 hour ago
OpenAI quietly raises 5-hour rate limits ~50% across plans
Kimmonismus·1 hour ago
Z.ai by email
Get an email when Z.ai has news
No news that day, no email.
Grok analyzes 47-minute AI employee transcript
Robert Scoble·1 hour agoEEBench measures whether AI can design circuit boards
EEBench uses atopile to test AI circuit design, avoiding GUI clicking. OpenAI's GPT-6 Astra demo in KiCad sparked the question. Models know electronics but real-world constraints like capacitor behavior remain challenging.
Hacker News·1 hour ago
Databricks achieves extreme efficiency via specialized GPU kernel generation
Databricks details a method for generating specialized GPU kernels to replace generic ones in production inference, aiming for extreme efficiency. The approach targets diverse workloads that generic kernels handle inefficiently.
Databricks Blog·1 hour ago

Benchmark of 21 Qwen3.8 27B variants on 16GB VRAM
A Reddit user benchmarked 21 Qwen3.8 27B variants on an RTX 5080 with 16GB VRAM, testing on C code. Best overall was bartowski/Qwen3.8-27B-IQ4_XS; some quants underperformed.
r/LocalLLaMA·1 hour agoLangSmith showcases agent observability with customer stories
LangChain·2 hours ago