Google ships voice AI that talks while it thinks

Google DeepMind launched Gemini 3.8 Live and Extended Thinking Sep 15 to make voice agents reason and use tools without breaking conversation.

Google Gemini 3.8 Live voice interaction on phone and laptop
Gemini 3.8 Live and Extended Thinking add real-time reasoning to voice· Deepmind

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking arrived Sep 15, 2026, as Google DeepMind’s most advanced live dialogue models, built to reason, see and call tools without pausing the conversation.

The pitch is simple: talk never stops.

According to Deepmind, the standard 3.8 Live is tuned for scale and cost efficiency with fluid dialogue and visual grounding, while Extended Thinking adds multi-step reasoning and speaks while it works, opening with cues like "Let me check that…" and narrating progress live. The team of Tom Ouyang, Principal Engineer, and Malini Jaganathan, Member of Technical Staff, put numbers on it: Extended Thinking took the #1 spot on Artificial Analysis’ Speech to Speech Quality Index at 82.6, hit 68.6% on τ-Voice and 35.1% on Sierra’s τ-Voice-banking, and scored 97.7% on Big Bench Audio. Gemini 3.8 Live took second place in the Speech Agent Arena and both pushed the Pareto Frontier on ServiceNow’s EVA-Bench for balancing accuracy with conversational quality.

What had to be true before this was already in motion. Google needed a Live API that could stream audio and video in near real time, handle 97 languages with mid-sentence switching, and run tool calls in the background. It also needed distribution: Gemini app, Google Workspace and Search Live where these models would actually live. The Sep 15 post was updated Sep 17, 2026, but the developer positioning was clarified off-blog, where Google lists Gemini 3.8 Live as the default option for most low-latency voice agent experiences and real-time dialogue without reasoning-induced delays.

What still has to happen is rollout. Both models are live now in the Gemini API and Google AI Studio. For enterprises they are in private preview in Gemini Enterprise, with Customer Experience and Google Workspace business access flagged as coming soon. For consumers, 3.8 Live is in Search Live and Extended Thinking is in Gemini Live for Google AI Pro and Ultra subscribers, plus Docs Live, Gmail Live and Keep Live. The demos point to that future: onboarding with live visual context, playing chess from a camera feed, turning sketches into React components on voice feedback, and chaining bookings through asynchronous function calls.

Google says all generated audio is watermarked with SynthID, woven imperceptibly into the output to help detection. That helps provenance, but only if detectors are deployed and checked, it does not stop a live voice agent from being misused in the moment.

© 2026 StartupHub.ai. All rights reserved. You may not republish this article in full without a license. Search engines and AI research tools may crawl and summarize for reference. Bulk reproduction or model training requires a license. See our terms.
Daniel Singer

Written by

Daniel Singer

Editor, StartupHub.ai

Daniel Singer is the editor of StartupHub.ai, a technology expert and thought leader on AI and its applications across sectors, from fintech and healthcare to developer tooling and consumer software. He writes and tests the tools covered here thoroughly and regularly, and built StartupHub.ai to give founders, operators and buyers a clearer read on what they are actually being sold.