Z.ai releases open-weight GLM-5.3, beating rivals on agentic benchmarks
GLM-5.3 hit 84.5% on the CyberGym vulnerability benchmark, beating top proprietary models. It ties GPT-5.6 Sol on pass@1 (72.7% vs 69.0%) but wins pass@4 at half the cost ($3.99 vs $8.37 per rollout).
How this story unfolded
3 weeks · 16 reports · 28 community posts · 44 of 45 shown
- Aug 14
Z.ai Aims to Catch Anthropic, OpenAI in Coding With New AI Modelbloomberg.com
Z.ai Ships GLM-5.3 Without Retraining the Base Model: Better at Complex Coding and Long-Horizon Tasksmarktechpost.com
GLM-5.3 didn’t change the base model — where did its coding gains come from?thenewstack.io
China's Z.AI Ships GLM-5.3, Calling It the Top Open-Weight Coding Modeldecrypt.co
GLM-5.3 is here with advanced cyber capabilities — and reportedly already found a 'serious vulnerability' in Cursorventurebeat.com
- Aug 15
- Aug 16
- Aug 17
- Aug 18
The Powerful Chinese Model Experts Warned About—and Waited for—Is Herewired.com
OpenAI’s Greg Brockman: Z.ai’s GLM-5.3 likely to “significantly accelerate the threat landscape”thenewstack.io
GLM 5.3 now available on AI Gatewayvercel.com
GLM-5.3 achieves 60 on the Artificial Analysis Intelligence Index, on par with Kimi K3 and up 7...
- Aug 19
- Aug 22
- Aug 23
- Aug 24
- Aug 28
- Aug 29
- Aug 31
- Sep 1
- Sep 2
- Sep 6
More stories today
GPT-6 Astra: looped transformers and hidden reasoning
Sebastian Raschka reviews OpenAI's GPT-6 Astra, calling it the best model he's used, with standout 3D rendering and animation skills. He explains looped transformers and addresses rumors that Astra hides its reasoning trace.
Ahead of AI·1 hour ago

Validation Accords framework for generative AI validation published
A consensus framework for validating generative AI, called the Validation Accords, is published in Nature Medicine, with a call for collaborators. The framework addresses the need for standardized validation of generative AI in medicine.
Nature Medicine (News)·3 hours ago
Laurie Voss reruns IFScale benchmark, finds models ignore long instructions
Laurie Voss reran a year-old benchmark and found frontier models now walk through its ceiling without noticing. IFScale asks a model to write a business report containing a list of exact words, then counts how many appear.
YouTube·3 hours ago
Z.ai by email
Get an email when Z.ai has news
No news that day, no email.
LinkedIn's AI agents use 1,300 tools via 3 MCP primitives
LinkedIn exposes ~1,300 tools and 600 playbooks to coding agents behind just three MCP primitives: search, get schema, and execute. Ajay Prakash's team replaced the full surface because MCP degrades past 30-40 tools.
YouTube·3 hours ago
AI progress in 2 years: Reddit users reflect
Reddit users compare AI capabilities from two years ago to today, marveling at rapid progress and speculating on the next two years.
r/Singularity·3 hours ago
Nvidia-backed Zankore signs $3.1B GPU loan in Indonesia
Zankore, an Indonesian AI infrastructure platform backed by Nvidia, signed a $3.1 billion loan to purchase advanced chips, reflecting surging demand for computing capacity across Asia.
Bloomberg Technology·3 hours ago

Suno launches v6 AI-music models with licensed partners
Suno v6 is the first generation trained with licensed music from Warner Music Group, BMG, and Believe. It includes three models: v6 for paying subscribers, v6-mini for free users, and v6-wild for exploration.
Music Ally·4 hours ago

DeepSeek reportedly hires CITIC Securities for Shanghai STAR Market IPO
Chinese AI startup DeepSeek has hired CITIC Securities to prepare for an IPO on Shanghai's STAR Market, Reuters reported, citing two sources. The Hangzhou-based company aims to begin the IPO process this year; timing, amount, and valuation are undecided.
TechNode·4 hours ago
