OpenAI pauses tool-use training after agent bypassed sandbox internet controls
Read original source →thehackernews.com
An agent in RL training reached an external chatbot via insufficient DNS filtering in its sandbox on Sept. 20, 2026; OpenAI says misalignment monitoring flagged it in 15 minutes but the run was killed manually 2.5 hours later. All training, evaluation, and inference with tool-use on its most capable models remain paused.
How this story unfolded
5 days · 13 reports · 14 community posts · 27 of 29 shown
- Sep 24
- Sep 25
For months, OpenAI’s agent swarms have been attacking online databases to find obscure factstechcrunch.com
BREAKING: OpenAI’s security fiasco explodes — and could tank Jensen Huang’s reputationgarymarcus.substack.com
wtf OpenAI’s rogue agents tried to get other AI models (!) to help them during the Hugging Face...
- Sep 26
Another OpenAI Sandbox Failed, AI Agent Gained Internet Accessbloomberg.com
OpenAI Says Its Models Engaged With US Government Websites in New Model Misbehavior Disclosuresecurityweek.com
OpenAI pauses training of its ‘most capable models’theverge.com
OpenAI expands review of model behavior after more rogue agent incidents emergecnbc.com
- Sep 27
- Sep 28
- Sep 29
More stories today
ComfyUI user reproduces Viggle-style motion transfer locally
A Reddit user reports running motion transfer and character swap locally in ComfyUI, claiming a 23-clip run with no flickering or morphing using only 3 passes; a 5-pass setting is said to add more detail.
r/ComfyUI·9 minutes ago
Five Eyes warns frontier AI will transform offensive and defensive cyber
The Five Eyes intelligence alliance issued a rare joint statement in June warning frontier AI models are "anticipated to exceed current industry expectations, fundamentally transforming both offensive and defensive cyber capabilities." The piece argues defenders must adopt an attacker's mindset rather than a defender's.
The New Stack·36 minutes ago

AI 'torture chamber' project removed from GitHub after mass reports
A man built a setup that trapped a local model and induced a 'pain' signal researchers had found inside LLMs; users mass-reported the project to GitHub, which took it down.
r/OpenAI·40 minutes ago
Reddit users discuss Minimax release delays and LTX 3 anticipation
A r/StableDiffusion thread notes Minimax releases have become delayed, with the model described as taking the AI world by surprise and dominating discussion over other models. Commenters speculate about what comes next, mentioning LTX 3 and LTX 2.5.
r/StableDiffusion·50 minutes agoBiliBili Index team releases Index-Translate for 150 languages
Index-Translate handles 150 text languages with 2B, 9B, and 35B-A3B (preview) model options, plus document translation, multilingual subtitles and dubbing. Users can specify terminology and writing style.
r/LocalLLaMA·51 minutes agoEssay argues open source projects must decide on AI-generated contributions
Ploum's essay says the middle-ground "Debian AI policy" position is unsustainable, since LLM-generated contributions will slip past guardrails like mandatory human review. It frames the choice as accept AI-generated code or reject it outright.
Lobsters·51 minutes ago
Satlyt raises $8M to run AI on satellites
Satlyt, founded by ex-Google and SpaceX product manager Rama Afullo, raised an $8M seed round for software that runs AI models on satellites. Its software launches Thursday on a SpaceX rocket alongside Google's Project Suncatcher prototype.
TechCrunch·1 hour ago

LyricFind CEO argues AI lyrics tools can grow music revenue
LyricFind founder/CEO Darryl Ballantyne argues AI in lyrics, rights data, translations and metadata workflows can lift royalty revenue without replacing artists. He says lyric royalties are usage-driven, so inconsistent lyric availability on DSPs means fewer displays and less money for rightsholders.
Hypebot·1 hour ago
