Google Launches Gemini 3.5 Transcribe, a 70% Faster Speech-to-Text Model Supporting 85+ Languages
Summary
Google launches Gemini 3.5 Transcribe, its fastest and most accurate speech-to-text model yet, boasting 70% faster transcription speeds, support for 85+ languages, and smart filler-word filtering — now powering Android's Rambler feature and available to developers via Google AI Studio, threatening to disrupt market leader Wispr Flow.
Key Points
- Google launches Gemini 3.5 Transcribe, its most precise speech-to-text model yet, powering the Rambler feature on Android and the Gemini app on macOS, with availability coming soon to Chrome.
- The model supports over 85 languages, filters out verbal corrections and filler words, recognizes speaker-specific jargon, and is available via two APIs — one for real-time interactions and one for pre-recorded audio — outperforming its predecessor by 70% on transcription speed benchmarks.
- By opening Gemini 3.5 Transcribe to developers through Google AI Studio, Google is poised to flood the market with Rambler-like experiences, putting significant competitive pressure on current market leader Wispr Flow.