Androidauthority iconAndroidauthorityAug 26, 2026 ~3 min source read

Google launches Gemini 3.5 Transcribe for more accurate, context-aware speech-to-text

Gemini 3.5 Transcribe replaces Google’s Chirp 3 transcription model and is rolling out across Search Live, Gemini Live, Docs, Keep, Gmail, the Gemini app, and Gboard.

Google rolls out Gemini Audio to improve real-time dialogue and speech recognition

Share this story

Send the public story page.

Useful takeaways from this story.

Gemini 3.5 Transcribe is Google’s newest speech-to-text model and replaces Chirp 3.

It offers automatic language detection for over 85 languages, removes filler words, supports custom vocabulary, and timestamps up to three speakers in pre-recorded audio.

Reported accuracy: 4% word-error-rate on streaming audio and 2.6% on pre-recorded files in varied real-world conditions.

# What Gemini 3.5 Transcribe is and where you'll see it

Google introduced Gemini 3.5 Transcribe as its latest speech-to-text model. It replaces the previous transcription engine, Chirp 3, and is designed to deliver cleaner, more accurate transcripts across live and pre-recorded audio. Google says the model is rolling out broadly: Search Live, Gemini Live, Docs, Keep, Gmail, the Gemini app, and Gboard will get the update. Developers will access it through the Gemini API in Google AI Studio and Antigravity, and enterprise customers via Gemini Enterprise offerings.

# Practical improvements you can expect

Gemini 3.5 Transcribe focuses on cleaner output and better context handling. Concrete features announced by Google include:

  • Automatic removal of filler words like "um" and "ah" so transcripts read more smoothly.
  • Auto-formatting and voice-driven editing so spoken corrections and edits can produce natural text with less manual cleanup.
  • Automatic language detection for more than 85 languages, plus improved handling of regional accents and dialects.
  • Custom vocabulary support to recognize specialized jargon and unconventional spellings you supply.
  • Function calling that lets the transcription model hand off tasks (for example, image generation) to other Gemini models.

# Accuracy and performance details

Google reported a 4% word-error-rate (WER) on streaming audio and 2.6% WER on pre-recorded files across diverse real-world conditions that include background noise and conversational AI interactions. Those values indicate lower error rates than the previous model according to Google's testing and are meant to reduce the time users spend correcting voice-to-text output.

# How Google is integrating the model

Some products already show the model in action. The Rambler feature in Gboard uses this transcription capability in select countries and languages. Google also plans Chrome integration so you can "talk to type" in any web field. Beyond consumer apps, the model will be made available to developers through Google AI Studio and Antigravity APIs, and to enterprise customers via Gemini Enterprise Agent Platform and Gemini Enterprise for Customer Experience.

# What this means for different users

  • Casual users: Expect cleaner voice typing in Gmail, Docs, Keep, and the Gemini app with fewer filler words to edit out.
  • Professionals and teams: Custom vocabulary and multi-speaker timestamps aim to make meeting notes and interviews more accurate and ready to share.
  • Enterprises: The model will be included in Gemini Enterprise products intended for customer experience and agent workflows.

# Short summary

Gemini 3.5 Transcribe targets faster, more accurate and more usable transcripts by removing filler words, recognizing specialized vocabulary, handling many languages and accents, and attributing speakers in recordings. Google is rolling it out across consumer apps, developer tools, and enterprise products.

More context around this story.

Amrc iconAmrcAug 7, 2026

What's happening

Find the latest AMRC news, details about meetings and events, and keep up to date with our newsletters and blogs. News Blog Events Newsletters Networks and groups

Loading more related stories...

Keep reading in the app

Open the app view to save this story, compare related coverage, and continue from the same source.

Open in app