Google launches Gemini 3.8 Flash, its third Flash model in six weeks

Gemini 3.8 Flash is available today at $0.75/1M input and $3.75/1M output tokens, the same introductory price as 3.7 Flash. It scores 73.7% on DeepSWE 1.1, outperforming competitors, but may use more tokens at higher effort levels.
How this story unfolded
3 days · 10 reports · 27 community posts · 37 of 41 shown
- Sep 1
- Sep 2
Gemini 3.8 Flash in Google Antigravityantigravity.google
Introducing Gemini 3.8 Flash and 3.8 Flash Cyberblog.google
Gemini 3.8 Flash now available on AI Gatewayvercel.com
Google ships its third Gemini Flash model in six weeksthenewstack.io
Google DeepMind Releases Gemini 3.8 Flash and Gemini 3.8 Flash Cyber: One Core Model, Two Access Envelopesmarktechpost.com
llm-gemini 0.34simonwillison.net
Google releases Gemini 3.8 Flash, its third Flash model in six weeksarstechnica.com
Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost moretheverge.com
- Sep 3
- Sep 4
More stories today
OpenAI claims AI solved Navier-Stokes problem; academics cry foul
OpenAI said it found an AI-generated solution to the Navier-Stokes equation, a Clay Millennium Prize problem worth $1 million. The claim is disputed by mathematician Tristan Buckmaster, who alleges OpenAI rushed ahead after learning of his and Levent Alpöge's progress.
Wired·58 minutes ago

Lab develops zero-downtime embedding model migration
A lab claims a method to migrate between embedding models with zero downtime, useful for RAG systems with large document corpora. Details are sparse, with no technical specifics or benchmarks provided.
r/MachineLearning·1 hour ago
AWS benchmarks small LLM inference on SageMaker AI: G7 vs G5 and G6
AWS compares GPU instances for small LLM inference, showing a generation jump can slash latency, increase throughput, and reduce cost-per-token. The post provides real-world benchmarks on SageMaker AI.
AWS AI Blog·1 hour ago

Google DeepMind by email
Get an email when Google DeepMind has news
No news that day, no email.
Google Cloud and Accenture launch joint AI deployment unit
Google Cloud and Accenture formed the Accenture Gemini Enterprise Business Group, a unit sending forward-deployed engineers into enterprises to boost AI adoption. Google will train up to 1,000 Accenture FDEs.
TechCrunch·1 hour ago

HPE Zerto builds agentic troubleshooting system with Amazon Bedrock
HPE Zerto built an agentic troubleshooting system using Amazon Bedrock to assess health, investigate issues, and act on problems across hybrid and multi-cloud infrastructures. The post was co-written by AWS and the HPE Zerto team.
AWS AI Blog·1 hour ago

DiDi builds contact center QA on Amazon Bedrock
DiDi's International Business Group built an intelligent contact center QA system on Amazon Bedrock, covering Spanish and Portuguese across ride-hailing, food delivery, and financial services.
AWS AI Blog·1 hour ago

ChatGPT suggests functional AGI may precede human-like AGI
A ChatGPT response posits that functional AGI may be achieved before human-like AGI, sparking discussion on r/ChatGPT with 67 comments.
r/ChatGPT·1 hour agoStudy: Data improvements drive 3.24x more compute efficiency gains than model tweaks
From 2019 to 2025, data improvements yielded 12.0x compute efficiency gains versus 3.7x from model improvements at 1e19 FLOPs, with gains mostly independent. Findings based on training year-representative recipes and data corpuses up to 1e19 FLOPs, evaluated on OLMES.
Dwarkesh Patel·1 hour ago
