Sakhanda Wire
NVDA MSFT GOOGL META AMZN
← Back to the news

Google announces Gemini 3.5 Transcribe for AI-powered speech-to-text

While we wait (possibly in vain) for Gemini 3.5 Pro to launch, Google is releasing a different model in the 3.5 branch. The company has announced Gemini 3.5 Transcribe, an AI model designed to streamline voice input by editing out "ums" and corrections, outputting polished AI text. This model already powers the Gboard "Rambler" feature on the Pixel 11, but it's about to appear throughout the Google ecosystem.

According to Google, Gemini 3.5 Transcribe is much faster and more accurate than its previous voice-to-text engine, known as Chirp 3. The new AI model should be about 70 percent faster from voice to final transcribed text, and the live-speech error rate has dropped to 5.5 percent. That's only a little better than Chirp 3, which Google measures at 7.32 percent. Still, it's a pain to fix typos when you're using voice input, so any improvement here is beneficial.

Credit: Google

Read full article

Comments

Originally published by Ars Technica on

Read the original on Ars Technica ↗

Text and images are the property of Ars Technica and are reproduced here with attribution and a link to the original publication.

← Back to the news

More stories

All the latest news