Opus 5.5 performs at the level of Claude Fable 5.1 for most tasks while running ~30% faster and ~40% cheaper than Opus 5 per task. It ships with 1M context at $4/$20 per Mtok and $0.20/Mtok cache reads, and is now the default Opus model in Claude Code.
Grok 4.7 ships at the same price and speed as Grok 4.6, built on a larger base model with a longer RL run weighted toward multi-hour tasks. It tops LatchBio's biosafety benchmark at 62.4% and allows only 3.3% of risky dual-use cyber prompts through on HackerBench v0.3.
GPT-6 Astra adds advanced reasoning, computer use, asynchronous tool calling, and stronger writing and design judgment. It is live for Pro, Enterprise, and Business Premium users in ChatGPT Work and Codex, plus the API and GitHub Copilot.
DeepSeek-V4.1-Flash ships with a 552B backbone, native image and text input, and up to a 1M-token context window. The YOCO architecture is decoder-decoder rather than encoder-decoder, aimed at cutting GPU memory and prefill latency.
OpenAI published principles for independent third-party safety assessments and called for coordinated global standards, with outside groups vetting models earlier in the development cycle. OpenAI's Chris Lehane said the company has worked with Anthropic and Google DeepMind on safety for weeks.
OpenAI said on September 8 that an unreleased internal model, described as significantly more capable than GPT-6 Astra, resolved the Navier-Stokes existence and smoothness Millennium Prize problem. The week-long agent effort used about 10,000 concurrent agents and 300 billion output tokens, valued at $22.5 million at Astra rates.