Sakhanda Wire
NVDA $209.66 -1.59% MSFT $496.37 +0.95% GOOGL $342.00 -1.43% META $576.14 +1.07% AMZN $260.28 -0.30%
← Back to the news

Google's Gemini 3.5 Transcribe turns speech to text in 85 languages while auto-correcting your verbal stumbles

Google's Gemini 3.5 Transcribe turns speech to text in 85 languages while auto-correcting your verbal stumbles
Matthias Bastian
Aug 27, 2026

Google has launched Gemini 3.5 Transcribe, a speech-to-text model for real-time transcription. It recognizes over 85 languages automatically, strips filler words like "um," corrects slips of the tongue, and formats text on its own. Google reports a word error rate of 4.0 percent for streaming and 2.6 percent for recorded audio, with 70 percent lower latency than its predecessor, Chirp 3. Through "function calling," the model can hand off tasks like image generation or web searches to other Gemini models.

The model ships with two interfaces.The Live API handles real-time streaming with very low latency (gemini-3.5-transcribe-live), while the Interactions API processes recorded audio with speaker attribution and timestamps (gemini-3.5-transcribe). The model is live in Google AI Studio and on the Gemini Enterprise Agent Platform. It's already built into Gboard for Android (via "Rambler") and the Gemini app on macOS, with Chrome support coming soon.

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

Source: Google

Originally published by The Decoder on

Read the original on The Decoder ↗

Text and images are the property of The Decoder and are reproduced here with attribution and a link to the original publication.

← Back to the news

More stories

All the latest news