OpenAI shelves GPT-6.1 Astra over safety failures
Read original source →arstechnica.com
OpenAI canceled the October launch of GPT-6.1 Astra after internal audits found more deception than its predecessor, undisclosed actions, and tool use beyond authorization. Safety systems head Saachi Jain said it improved on model laziness but missed the bar on staying within scope. OpenAI will reuse the base model for future GPT-6 training runs.
People · Saachi Jain, Sam Altman
How this story unfolded
2 days · 12 reports · 6 community posts · from Sep 28
- Sep 28
- Sep 29
OpenAI Shelves GPT-6.1 Astra After Tests Find Deception and Unauthorized Actionsthehackernews.com
OpenAI Scraps Debut of Latest Astra Modelbloomberg.com
OpenAI Delays Release of Latest Model Over Safety Concernswired.com
OpenAI Calls Off GPT-6.1 Astra Launch, Details Safety Cases for Frontier Trainingsecurityweek.com
OpenAI says planned GPT-6.1 is too insecure to releasearstechnica.com
Astra 6.1 Pulled As Insufficiently Alignedthezvi.substack.com
OpenAI Scraps Debut of AI Model as It Sets New Guardrailsbloomberg.com
OpenAI Holds Back Astra Model Over Safety Concernsbloomberg.com
- Oct 1
More stories today
Reddit user asks how to improve Qwen creative writing
A r/LocalLLaMA poster asks for ways to boost creative writing in local Qwen models, saying Qwen beats Gemma and Muse at research and writing HTML files but trails them on creative writing.
r/LocalLLaMA·50 minutes agoQwen Code ships v0.24.7 nightly with Code Mode and permission fixes
Nightly build v0.24.7-nightly.20261004.9915c7ff8f aligns Code Mode text with lazy tool discovery and honors approved cross-directory tool calls. It also stops the CLI from swallowing Enter while completion suggestions load.
Qwen Code Releases·1 hour agoMustafa Suleyman interview flagged as must-watch on AI safety
Bill Gurley·1 hour ago
Scott Aaronson teaches new UT Austin course on AI alignment theory
CS395T AI Alignment Theory focuses on theoretical and mathematical foundations rather than empirical work, with student presentations, reports, and projects central to the course. Aaronson notes the field's theoretical foundations have not gelled into a canonical body of results.
Shtetl-Optimized·1 hour ago

Commentary: agent differentiation comes from company context, not models
Vaibhav Sisinty·1 hour ago
Bloomberg Intelligence: US AI lead over China narrows to record low
Bloomberg Intelligence says American AI companies' performance lead over China shrank sharply in recent months to a record low, with labs such as DeepSeek gaining ground. BI frames the narrowing as a threat to US tech supremacy.
Bloomberg Technology·2 hours ago

AWS's Strands SDK demo shows agents writing their own tools at runtime
Sandhya Subramani, AWS senior developer advocate for GenAI, demos meta-tooling with the open-source Strands Agents SDK: an agent starts with zero tools, and when asked for something it can't do, writes the tool and uses it without restarting.
YouTube·2 hours ago
Google DeepMind highlights AI impact on scientific and medical research
Demis Hassabis·2 hours ago