Marktechpost iconMarktechpostSep 15, 2026

Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents

Google has released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its most advanced live dialogue models to date. The models execute tools and API calls in the background while the conversation keeps flowing, process live visual inputs, and switch between 97 languages mid conversation.

Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents

Share this story

Send the public story page.

Useful takeaways from this story.

Google has released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its most advanced live dialogue models to date.

The models execute tools and API calls in the background while the conversation keeps flowing, process live visual inputs, and switch between 97 languages mid conversation.

Both are available today in the Gemini API and Google AI Studio at $0.005/min for audio input, with all generated audio carrying Google DeepMind's SynthID watermark.

Building the complete brief

The page is ready to read now. The fuller skim-friendly version will appear here automatically.

The useful part

Google has released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its most advanced live dialogue models to date. The models execute tools and API calls in the background while the conversation keeps flowing, process live visual inputs, and switch between 97 languages mid conversation. Extended Thinking ranks #1 on Artificial Analysis' Speech to Speech Quality Index with 82.6 and scores 97.7% on Big Bench Audio.

How it works

  • Both are available today in the Gemini API and Google AI Studio at $0.005/min for audio input, with all generated audio carrying Google DeepMind's SynthID watermark.

Keep reading in the app

Open the app view to save this story, compare related coverage, and continue from the same source.

Open in app