Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Google DeepMind announces Gemini 3.5 Transcribe, a new speech-to-text model for real-time transcription, available via the Gemini API in Google AI Studio and Gemini Enterprise Agent Platform. The model offers smart transcription, function calling, custom vocabulary, multilingual support, and multi-speaker identification. It is accessible through streaming (Live API with gemini-3.5-transcribe-live) and pre-recorded audio (Interactions API with gemini-3.5-transcribe) APIs.
From the source
Today, we’re introducing Gemini 3.5 Transcribe, our most precise speech-to-text model yet, designed for intelligent voice interactions.
deepmind.google