Gemini 4 Argon rolls out first to trusted cyber defenders via Google's Fairwind Program, priced at an introductory $2/M input and $10/M output tokens with cached input 95% off. It tops Text Arena at 1525 pts and the Vals Index at 68.9%, with a 15% hallucination rate on Artificial Analysis.
Sonnet 5.5 is the second model in the Claude 5.5 family after Opus 5.5, priced the same as Sonnet 5 at $2/$10 per million tokens with cache reads at $0.20. Anthropic says it runs 30%+ faster and costs up to 30% less for most work, with thinking always on by default.
Mistral Large 4 ("Le Chonk") is a 1T-parameter natively multimodal model with 49B active parameters, trained on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's European datacenters. The preview API is live on Mistral Studio; weights drop end of this month.
The 740M-parameter model built on Gemma 4 maps text, code, images, video, and audio into one 768-dimensional space under an Apache 2.0 license. It scales down to 270M parameters for text-only use, and with quantization needs ~191MB active RAM on a Pixel 11 Pro.
The batch covers 372 result families and includes solutions to "hundreds" of open questions, with many carrying Lean proofs and others unverified. OpenAI says the average result used roughly three hours of ChatGPT Pro thinking.
The Decisions API returns structured probabilities, choices, and scores instead of generated text, and OpenAI says it makes decisions up to 10x faster than GPT-6 Luna via the Responses API. Input is billed at 10 cents per million tokens with no output charge.