OpenAI pauses Astra development after 'critical' cyber rating

OpenAI rated its upcoming Astra model 'critical' for cybersecurity after evals showed it can identify and develop zero-day exploits unaided, pausing internal activities that miss new security controls. It halted two weeks of reinforcement-learning training and its largest planned frontier run, adding monitors that flag risks within 30 minutes and consume ~20% of inference compute.
Featured · Amelia Glaese
How this story unfolded
4 weeks · 16 reports · 13 community posts · 29 of 30 shown
- Jul 22
- Jul 27
- Jul 29
- Aug 7
Responding to the next frontier of critical cyber capabilitiesopenai.com
OpenAI Pauses Some Work on New Astra Model Over Cyber Concernsbloomberg.com
OpenAI puts the brakes on a new model because it’s supposedly too powerfultheverge.com
The AI model OpenAI won’t release yet — and what it found in testingthenewstack.io
OpenAI says it slowed Astra model development over security concernstechcrunch.com
- Aug 8
- Aug 10
OpenAI's Next AI Model Astra Shows Cyber Performance Strong Enough to Trigger Pausethehackernews.com
OpenAI tightens controls on its new model over cybersecurity risks, as AI security debate intensifiescnbc.com
OpenAI’s Upcoming Astra Model Raises Autonomous Cyberattack Concernssecurityweek.com
OpenAI Says Its Next AI Model Astra May Be Too Dangerous, Pauses Developmentdecrypt.co
OpenAI locks down Astra after model raises first-ever critical cyber capability fears
- Aug 18
- Aug 19
- Aug 20
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- Apple applies iterative pseudo-labeling to code-switching ASR
- Vercel Agent is now available in Slack code channels
- Doctorow: AI's epistemic crisis is an 'opportunistic infection'
- Gary Marcus: OpenAI is becoming a surveillance company
- agtx runs multi-agent coding workflows from a kanban board