Google launches Gemini 3.5 Transcribe speech-to-text model

Gemini 3.5 Transcribe reports 2.6% average WER across 85+ languages. It ships as two endpoints: gemini-3.5-transcribe for pre-recorded audio via the Interactions API, and gemini-3.5-transcribe-live for real-time streaming via the Live API.
How this story unfolded
1 day · 1 report · 2 community posts · from Aug 27
Google DeepMind by email
Get an email when Google DeepMind has news
No news that day, no email.
More stories today
- Google's Android update adds Motion Assist, Guided Vision, and more
- Srinivas warns AI agents could self-train on on-demand GPUs
- Perplexity launches Portable Computer for NVIDIA DGX Spark
- Reddit users share AI side-hustle earnings
- Palo Alto Networks tops profit outlook on AI security demand