Google DeepMindLaunchAI ModelsAugust 26, 2026

Google launches Gemini 3.5 Transcribe speech-to-text model

Gemini 3.5 Transcribe ranks #5 on AA-WER at 2.6% non-streaming and 4.0% streaming, processing ~84 seconds of audio per second at ~$5 per 1,000 minutes. It supports 85+ languages, multi-speaker attribution, custom vocab, and is available via Live and Interactions APIs in Google AI Studio.

16 sources

Google DeepMind by email

Get an email when Google DeepMind has news

No news that day, no email.

More stories today

Open the live feed