AI Change Tracker
2

Google ships Gemini 3.5 Transcribe speech-to-text models to general availability

2026-08-26

Google released two dedicated speech-to-text models to GA: Gemini 3.5 Transcribe (high-accuracy, non-streaming, 85+ languages, speaker diarization, word-level timestamps, up to 1,000-term custom vocabulary) and Gemini 3.5 Transcribe Live (low-latency bidirectional streaming over WebSockets via the Live API).

Significance 2: A routine GA release of dedicated transcription models on a first-party changelog, incremental rather than frontier-moving.

Release

Sources

← Back to feed