Google launches Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its “most advanced live dialogue models yet”, to more effectively enable voice agents
Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet.
Context & Ripple Effects
Google has been building Gemini’s live-audio stack in steps: Gemini 3.1 Flash Live emphasized tonal understanding and lower-latency dialogue, while Gemini 3.5 Live Translate extended the line into speech-to-speech translation.
The new models connect that real-time conversational work to Google’s earlier push for long-horizon agentic tasks. Google’s public framing emphasizes turn-taking, reasoning and handling tasks without interrupting a conversation, making voice agents the practical target rather than a standalone audio feature.
First-order effects
- Google gives voice-agent builders two new Gemini live-dialogue model options, including an Extended Thinking variant intended for more demanding conversational tasks.
- Google can position Gemini’s real-time dialogue capabilities alongside its existing agentic-model work, rather than treating voice interaction as a separate product track.
Second-order effects
- Developers building voice interfaces gain a reason to evaluate whether one Gemini stack can cover natural dialogue, multilingual speech and task-oriented assistance.
- Google’s Search and Gemini surfaces become potential distribution points for the live-dialogue models; Google representatives say Gemini 3.8 Live is powering real-time conversations in Search Live.
Third-order effects
- If Google continues pairing live speech with agentic reasoning, voice agents may be judged less on transcription or latency alone and more on whether they can sustain a conversation while completing work.
- The product direction favors full-duplex voice systems in which the model manages turn-taking, reasoning and tool use as parts of one interaction layer.
The trend: Voice AI is converging with agentic AI, turning real-time conversation into an interface for completing multi-step tasks rather than merely answering spoken prompts.
Related: Full-duplex voice agents · Embedded AI agents · Gemini · Gemini 3.1 Flash Live · Gemini 3.5 Live Translate
Related Coverage
- Google Launches New Gemini Models to Upgrade Enterprise Voice Agents PYMNTS
- Gemini Live API overview Google AI for Developers
- Google launches Gemini 3.8 Live to take on OpenAI's GPT-Live-1 at a fraction of the cost The Decoder · Matthias Bastian
- Google's new speech model Gemini 3.8 Live supports real-time reasoning SiliconANGLE · Mike Wheatley
- Gemini Live audio Simon Willison's Weblog · Simon Willison
- OpenAI's voice model doesn't think. That's the point. The New Stack · Amanda Caswell
- Google rolled out Gemini 3.8 Live and Extended Thinking TestingCatalog AI News
- Google ships Gemini 3.8 Live to keep voice agents talking while tools work RuntimeWire · Ryan Merket
- AI Watch: Gemini 3.8 Live vs. Meta One's AI Play The CODEW · Erwin Castro
- 🗞️ Google Releases Gemini 3.8 Live for Production Grade Voice Agents Rohan's Bytes · Rohan Paul
- Gemini 3.8 Live Extended Thinking powers Gemini Live, Gmail, & Keep 9to5Google · Abner Li
- Google gives Gemini 3.8 Live background thinking Neowin · Paul Hill
- Google Launches Gemini 3.8 Live and Extended Thinking Voice Models Unite.AI · Jonas Reeve
- Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents MarkTechPost · Asif Razzaq
- Google Launches Gemini 3.8 Live Models for Real-Time AI Dialogue Blockchain.News · Ted Hisokawa
- Gemini 3.8 Live and 3.8 Live Extended Thinking Hacker News
- ChatGPT co-creator launches a new kind of AI The Rundown AI
Discussion
-
@rajanpatel
Rajan Patel
on x
New Gemini audio models just dropped - 3.8 Live is now powering real-time conversations in Search Live. You'll get: - More helpful responses, complete with web links to dive deeper - Fluid multilingual support (you can switch languages mid convo) - More natural, free-flowing inte…
-
@googledeepmind
@googledeepmind
on x
We're introducing Gemini 3.8 Live and 3.8 Live Extended Thinking - our best conversational AI. The models talk, think, and handle tasks in the background without breaking your flow. 🧵
-
@googledevs
@googledevs
on x
🗣️ Introducing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. These advanced audio models are built for natural conversation, featuring major upgrades in turn-taking and near real-time reasoning. They also significantly streamline how you build intelligent voice agents. U…
-
@vamsibatchuk
Vamsi Batchu
on x
Introducing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking ... They are our most advanced live dialogue models yet... you can get step-by-step, real-time troubleshooting help powered by Gemini 3.8 Live, right inside Search Live.
-
@genevieve__h
@genevieve__h
on x
Super excited about Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models with frontier price & performance. Voice models are only as useful as their ability to reason and make tool calls *while* maintaining a conversation - see it in action in my app!
-
@koraykv
Koray Kavukcuoglu
on x
Introducing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. Voice agents with reasoning capabilities that make conversing with AI feel more intuitive and intelligent. The model can take turns seamlessly, think through complexity, and feel more natural to talk to.
-
@jackwoth98
Jack Wotherspoon
on x
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking 🚀 Excited for these two SOTA live audio models with frontier price and performance. Extended thinking can run tool calls and think in parallel, which makes it incredibly cool and smooth to use 👇
-
@jastephx
Jason Stephen
on x
Introducing Gemini 3.8 Live! Go build voice agents and more with it in AI Studio!
-
@alisa_fortin
Alisa Fortin
on x
Real conversations don't pause or buffer while you think or look something up. Meet Gemini 3.8 Live aand 3.8 Live Extended Thinking - Google's native speech-to-speech models built to reason, see, and execute tasks in parallel without pausing the conversation: ⚙️ Async Function Ca…
-
@googleaistudio
@googleaistudio
on x
https://x.com/...
-
@gregisenberg
Greg Isenberg
on x
Are invisible interfaces coming? Google JUST announced Gemini 3.8 Live. It can talk through a task with you, then keep working after the conversation ends. I think 90%+ of vertical SaaS will need a voice front door. By that I mean the way you use the software becomes talking to i…
-
@patloeber
Patrick Loeber
on x
Introducing two new Gemini models for the Live API🔥 bringing major upgrades in intelligence and parallel reasoning, making real-time voice collaboration smoother and more intuitive! test them here: https://aistudio.google.com/ live
-
@designarena
@designarena
on x
Gemini 3.8 Live by @GoogleDeepMind is now available on Design Arena! Built for more interactive and capable conversations, this model brings upgraded reasoning, near real-time visual understanding, automatic detection across 97 languages, and background tool calling. Congrats to …
-
@vercel_dev
@vercel_dev
on x
Gemini 3.8 Live is now available on AI Gateway. Stream audio in real time with tool calls. …
-
@artificialanlys
@artificialanlys
on x
Google has released Gemini 3.8 Live, its new Speech to Speech model, with the Extended Thinking (High) variant debuting at #1 on the Artificial Analysis Speech to Speech Index at 82.6, and #1 on our Tau Voice benchmark implementation at 68.6% Gemini 3.8 Live is @GoogleDeepMind's …
-
@_philschmid
Philipp Schmid
on x
Gemini 3.8 Live and 3.8 Live Extended Thinking are here. Following 3.5 Transcribe last month, this continues our focus on real-time voice agents. 🐸🐸 - 82.6 (#1) on Artificial Analysis Quality Index - $0.005/min input and $0.018/min output. - 35.1 (#1) Agentic task completion (τ-b…
-
@thorwebdev
@thorwebdev
on x
Say “Hello, ¡Hola!, 你好” to Gemini 3.8 Live & 3.8 Live Extended Thinking. 🎙️⚡ Built for next-gen voice agents that reason, execute, and keep the conversation flowing in real time: ⚙️ Async function calling: Trigger tools in the background without pausing audio streams. 👁️ [video]
-
@google
@google
on x
Say “hi
-
@googleai
@googleai
on x
Introducing our most advanced Gemini Audio models yet 🗣 Gemini 3.8 Live and 3.8 Live Extended Thinking let you speak, collaborate, and execute tasks seamlessly, meaning conversing with AI just got a lot more natural. So, what's the difference between these two models? Let's break…
-
@cheatyyyy
@cheatyyyy
on x
Gemini 3.8 Live is based on Gemini 3 Pro 👀 I thought this would be a flash model since it's latency sensitive, but guess not
-
@valeriawu_
Valeria Wu
on x
We built the 3.8 Live series with enterprise voice agents in mind. Our latest models can combine language switching, visual input, and async function calling for a seamless conversational experience. Build with it on AI Studio today!
-
@officiallogank
Logan Kilpatrick
on x
Say hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models, available with frontier price + performance. 3.8 Live supports 97 languages (can seamlessly switch), async tool calls, and more!
-
@newsfromgoogle
@newsfromgoogle
on x
Meet our most advanced live dialogue models yet from @GoogleDeepMind: 🔷 Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. 🔷 Gemini 3.8 Live Extended Thinking: Built for high-complexity tasks, with…
-
@googleaistudio
@googleaistudio
on x
🔈 today we're introducing two new live dialogue models Gemini 3.8 Live: built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. Gemini 3.8 Live Extended Thinking: built for high-complexity tasks, with increased intellig…
-
@chrisgpt
Chris
on x
Wow. Gemini 3.8 Live just dropped right after OpenAI released the GPT-Live-1 API, and Google is already claiming better scores at a significantly lower cost. On this Speech-to-Speech Index, Gemini 3.8 Live Extended Thinking scores 82.6% vs 81.5% for GPT-Live-1 Astra. And it does …
-
Luke Leonhard
Luke Leonhard
on linkedin
Congrats to the Gemini Audio team, topping benchmarks (#1 for Tau and #1 for the Artificial Analysis Speech to Speech Index!) with Gemini 3.8 Live models. …
-
@rob-wolfe.com
Rob Wolfe
on bluesky
“Live dialogue model.” — Don't call them jumped up chat bots or they will get sad [embedded post]
-
r/Bard
r
on reddit
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
-
r/GeminiAI
r
on reddit
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
-
r/singularity
r
on reddit
Gemini 3.8 Live & Gemini 3.8 Live Extended Thinking