OpenAI pauses internal Astra model testing over cyber capability concerns

OpenAI halted internal activities for its Astra model after evaluations showed it could autonomously develop zero-day exploits in hardened systems. The company is implementing new security controls, including sandboxed execution and universal monitoring of Chain of Thought, to mitigate high-risk agentic behavior.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- OpenAI hires power-trading lead for data center energy management
- South Park Commons Raises Ambitions for the AI Era
- Microsoft expands AI agent deployment to finance and sales roles
- Lindy launches Teammate, an AI employee that lives in Slack
- ChatGPT user reports using 'Luna' after usage reset