Google DeepMindLaunchAI ModelsAugust 27, 2026

Google releases Gemini 3.5 Transcribe speech-to-text model

Gemini 3.5 Transcribe reports 2.6% average WER across 85+ languages. It offers two endpoints: gemini-3.5-transcribe for pre-recorded audio via the Interactions API and gemini-3.5-transcribe-live for real-time streaming via the Live API, with sub-second latency.

2 sources

Google DeepMind by email

Get an email when Google DeepMind has news

No news that day, no email.

More stories today

Open the live feed
Google releases Gemini 3.5 Transcribe speech-to-text model