Google releases Gemini 3.5 Transcribe speech-to-text model

Gemini 3.5 Transcribe reports 2.6% average WER across 85+ languages. It offers two endpoints: gemini-3.5-transcribe for pre-recorded audio via the Interactions API and gemini-3.5-transcribe-live for real-time streaming via the Live API.
2 sources
Google DeepMind by email
Get an email when Google DeepMind has news
No news that day, no email.
More stories today
- Data center spending to hit $31.6T by 2050 on AI boom
- OpenAI explores hiding model 'thinking', raising safety concerns
- Emad Mostaque: Frontier models will one-shot at 10k tokens/sec
- LangSmith adds Messages View for agent debugging
- Notebook collection covers 30 LLM agent memory techniques