Skip to main content

Gemini 3.5 Transcribe generally available (GA)

By Steven Van ·

: Released two dedicated speech-to-text models based on Gemini's audio understanding: Gemini 3.5 Transcribe (gemini-3.5-transcribe): High-accuracy…

Google Gemini announced Gemini 3.5 Transcribe generally available (GA) on August 26, 2026.

What changed

: Released two dedicated speech-to-text models based on Gemini's audio understanding: Gemini 3.5 Transcribe (gemini-3.5-transcribe): High-accuracy, low-latency non-streaming speech-to-text with utterance-based language detection across 85+ languages, speaker diarization, word-level timestamps, and custom vocabulary biasing (up to 1,000 terms). Gemini 3.5 Transcribe Live (gemini-3.5-transcribe-live): Low-latency, bidirectional streaming speech-to-text over WebSockets using the Live API, supporting interim and finalized transcription events, Smart transcription mode, and multiple Voice Activity Detection (VAD) strategies. To get started, see the Audio transcription guide, the Live transcription guide, and the Gemini 3.5 Transcribe model page.

Google Gemini
Google Gemini
Google's AI assistant — write, plan, brainstorm, generate images, and analyze files with one of the most powerful multimodal models.
View Google Gemini →

Read the original announcement →

Read Gemini 3.5 Transcribe generally available (GA) on Creators Toolbox