Intelligent transcription with Gemini 3.5 Transcribe
Aug 26, 2026
Our latest speech-to-text model designed for precise and intelligent real-time transcription.
Diego Melendo Casado
Senior Director, Engineering, Gemini Audio
Luke Leonhard
Chief of Staff, Gemini Audio, on behalf of Gemini Audio Team
Listen to article
[[duration]] minutes
This content is generated by Google AI. Generative AI is experimental
Today, we’re introducing Gemini 3.5 Transcribe, our most precise speech-to-text model yet, designed for intelligent voice interactions. Unlike conventional speech recognition models that struggle with background noise, complex jargon, and disfluency cleanup, Gemini 3.5 Transcribe converts raw audio directly into accurate, polished, formatted text.
Across our products like the Gemini app and on Android, we’ve seen consumers already benefiting from this transcription model with new voice capabilities like Rambler on Android and in the Gemini app on macOS. Now, developers can build similar capabilities with Gemini 3.5 Transcribe in the Gemini API in Google AI Studio and Gemini Enterprise...
Copyright of this story solely belongs to blog.google. To see the full text click HERE