OpenAI's Astra model raises autonomous cyberattack concerns

OpenAI flagged its upcoming Astra model as potentially reaching a 'critical' cybersecurity risk threshold, surpassing GPT-5.6-Sol's 'high' rating. The company paused internal development lacking new security controls and deployed monitoring to intercept high-risk behavior.
1 source
Policy by email
Get an email when there's news on Policy
No news that day, no email.
More stories today
- Data center spending to hit $31.6T by 2050 on AI boom
- OpenAI explores hiding model 'thinking', raising safety concerns
- Emad Mostaque: Frontier models will one-shot at 10k tokens/sec
- LangSmith adds Messages View for agent debugging
- Notebook collection covers 30 LLM agent memory techniques