2
Google ships Gemini 3.5 Transcribe speech-to-text models to general availability
2026-08-26
Google released two dedicated speech-to-text models to GA: Gemini 3.5 Transcribe (high-accuracy, non-streaming, 85+ languages, speaker diarization, word-level timestamps, up to 1,000-term custom vocabulary) and Gemini 3.5 Transcribe Live (low-latency bidirectional streaming over WebSockets via the Live API).
Significance 2: A routine GA release of dedicated transcription models on a first-party changelog, incremental rather than frontier-moving.
Release
Sources
- primary Gemini API release notes — Gemini 3.5 Transcribe GA retrieved 2026-08-28