AnalysisPolicyJuly 18, 2026
OpenAI's GPT Red finds vulnerabilities 84% of the time via self-play

OpenAI's GPT Red uses self-play reinforcement learning to identify model vulnerabilities, finding 84% compared to 13% for human testers. The approach is part of a recursive self-improvement safety loop designed to catch weaknesses before deployment.
1 source
More stories today
- LogoCreator generates logos with Flux Pro 1.1 on Together AI
- Flank launches Record, an agentic contract truth system
- Gigatoken: Rust BPE tokenizer encodes text at 24.
- AI leader says 3D, video generation not on main path
- DeepSeek prioritizes AGI research over products and commercial growth