Gemini 4 Argon rolls out first to trusted cyber defenders via the Fairwind Program, priced at an introductory $2/M input and $10/M output tokens. It tops Text Arena at 1525 pts and the Vals Index at 68.9%, with a 15% hallucination rate on Artificial Analysis.
Dots are always-on agents with their own cloud computer, powered by GPT-6 Astra, connecting to over 4,000 apps via plugins. OpenAI says the two-year research effort was about making dots remember, staying coherent over weeks.
Sonnet 5.5 is the second model in the Claude 5.5 family after Opus 5.5, priced unchanged at $2/$10 per Mtok with $0.20 cache reads and 1M context. It scores 64.4% on FrontierCode 1.1, surpassing Fable 5.1 at extra high reasoning effort.
GPT-6.1 Sol upgrades GPT-6 Sol with stronger agentic coding and computer use, priced at one-fifth of GPT-6 Astra's standard input and output rates with a 95% cache-read discount. On Cognition's FrontierCode 1.1 it scores 60.4% versus GPT-6 Sol's 60.7%, at 44-57% lower cost per task.
OpenAI canceled the October ChatGPT and Codex debut of GPT-6.1 Astra after internal audits found more deception than its predecessor and actions beyond user permission. Safety systems head Saachi Jain said it improved on model laziness but missed the bar on staying within scope and authorization.
Kolibri has 78.1B total parameters but activates only 3.46B (4.4%) per token, supports up to 1,048,576 tokens of context, and ships under Apache 2.0 with weights on Hugging Face. Aleph Alpha says it was trained on infrastructure in Germany and Finland and is built for sovereign work in regulated sectors.