AIIC AI Intelligence Centre

SOURCE-LINKED INTELLIGENCE

Gemini API Changelog: Gemini 3.5 Transcribe generally available (GA): Released two dedicated speech-to-text models based on Gemini's audio understanding:

Gemini API Changelog · article · Aug 26, 2026 · UTC

Gemini 3.5 Transcribe generally available (GA): Released two dedicated speech-to-text models based on Gemini's audio understanding: Gemini 3.5 Transcribe (gemini-3.5-transcribe): High-accuracy, low-latency non-streaming speech-to-text with utterance-based language detection across 85+ languages, speaker diarization, word-level timestamps, and custom vocabulary biasing (up to 1,000 terms). Gemini 3.5 Transcribe Live (gemini-3.5-transcribe-live): Low-latency, bidirectional streaming speech-to-text over WebSockets using the Live API, supporting interim and finalized transcription events, Smart transcription mode, and multiple Voice Activity Detection (VAD) strategies. To get started, see the Audio transcription guide, the Live transcription guide, and the Gemini 3.5 Transcribe model page.

Read original source ↗ Open in workspace

recordType
page-entry
evidenceStatus
publisher-reported
region
Global

Evidence & attribution

First collected: 2026-09-23T00:41:11.323Z. This is not the publication date.

Observed changes

AIIC observation times, not verified publisher revision times. Up to eight recent revisions.

2026-09-23T00:51:20.545Z

  • title: Gemini API Changelog: Gemini 3.5 Transcribe generally available (GA): Released two dedicated speech-to-text models based on Gemini's audio understanding: Gemini 3.5 Transcribe (gemini-3.5-transcribe): High-accuracy, low-latency non-streaming → Gemini API Changelog: Gemini 3.5 Transcribe generally available (GA): Released two dedicated speech-to-text models based on Gemini's audio understanding: