/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

← → days · ↑ ↓ browse · Enter similar · o open

Google launches Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its “most advanced live dialogue models yet”, to more effectively enable voice agents

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet.

Google

Context & Ripple Effects

Google has been building Gemini’s live-audio stack in steps: Gemini 3.1 Flash Live emphasized tonal understanding and lower-latency dialogue, while Gemini 3.5 Live Translate extended the line into speech-to-speech translation.

The new models connect that real-time conversational work to Google’s earlier push for long-horizon agentic tasks. Google’s public framing emphasizes turn-taking, reasoning and handling tasks without interrupting a conversation, making voice agents the practical target rather than a standalone audio feature.

First-order effects

  • Google gives voice-agent builders two new Gemini live-dialogue model options, including an Extended Thinking variant intended for more demanding conversational tasks.
  • Google can position Gemini’s real-time dialogue capabilities alongside its existing agentic-model work, rather than treating voice interaction as a separate product track.

Second-order effects

  • Developers building voice interfaces gain a reason to evaluate whether one Gemini stack can cover natural dialogue, multilingual speech and task-oriented assistance.
  • Google’s Search and Gemini surfaces become potential distribution points for the live-dialogue models; Google representatives say Gemini 3.8 Live is powering real-time conversations in Search Live.

Third-order effects

  • If Google continues pairing live speech with agentic reasoning, voice agents may be judged less on transcription or latency alone and more on whether they can sustain a conversation while completing work.
  • The product direction favors full-duplex voice systems in which the model manages turn-taking, reasoning and tool use as parts of one interaction layer.

The trend: Voice AI is converging with agentic AI, turning real-time conversation into an interface for completing multi-step tasks rather than merely answering spoken prompts.

Discussion

  • @rajanpatel Rajan Patel on x
    New Gemini audio models just dropped - 3.8 Live is now powering real-time conversations in Search Live. You'll get: - More helpful responses, complete with web links to dive deeper - Fluid multilingual support (you can switch languages mid convo) - More natural, free-flowing inte…
  • @googledeepmind @googledeepmind on x
    We're introducing Gemini 3.8 Live and 3.8 Live Extended Thinking - our best conversational AI. The models talk, think, and handle tasks in the background without breaking your flow. 🧵
  • @googledevs @googledevs on x
    🗣️ Introducing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. These advanced audio models are built for natural conversation, featuring major upgrades in turn-taking and near real-time reasoning. They also significantly streamline how you build intelligent voice agents. U…
  • @vamsibatchuk Vamsi Batchu on x
    Introducing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking ... They are our most advanced live dialogue models yet... you can get step-by-step, real-time troubleshooting help powered by Gemini 3.8 Live, right inside Search Live.
  • @genevieve__h @genevieve__h on x
    Super excited about Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models with frontier price & performance. Voice models are only as useful as their ability to reason and make tool calls *while* maintaining a conversation - see it in action in my app!
  • @koraykv Koray Kavukcuoglu on x
    Introducing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. Voice agents with reasoning capabilities that make conversing with AI feel more intuitive and intelligent. The model can take turns seamlessly, think through complexity, and feel more natural to talk to.
  • @jackwoth98 Jack Wotherspoon on x
    Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking 🚀 Excited for these two SOTA live audio models with frontier price and performance. Extended thinking can run tool calls and think in parallel, which makes it incredibly cool and smooth to use 👇
  • @jastephx Jason Stephen on x
    Introducing Gemini 3.8 Live! Go build voice agents and more with it in AI Studio!
  • @alisa_fortin Alisa Fortin on x
    Real conversations don't pause or buffer while you think or look something up. Meet Gemini 3.8 Live aand 3.8 Live Extended Thinking - Google's native speech-to-speech models built to reason, see, and execute tasks in parallel without pausing the conversation: ⚙️ Async Function Ca…
  • @googleaistudio @googleaistudio on x
    https://x.com/...
  • @gregisenberg Greg Isenberg on x
    Are invisible interfaces coming? Google JUST announced Gemini 3.8 Live. It can talk through a task with you, then keep working after the conversation ends. I think 90%+ of vertical SaaS will need a voice front door. By that I mean the way you use the software becomes talking to i…
  • @patloeber Patrick Loeber on x
    Introducing two new Gemini models for the Live API🔥 bringing major upgrades in intelligence and parallel reasoning, making real-time voice collaboration smoother and more intuitive! test them here: https://aistudio.google.com/ live
  • @designarena @designarena on x
    Gemini 3.8 Live by @GoogleDeepMind is now available on Design Arena! Built for more interactive and capable conversations, this model brings upgraded reasoning, near real-time visual understanding, automatic detection across 97 languages, and background tool calling. Congrats to …
  • @vercel_dev @vercel_dev on x
    Gemini 3.8 Live is now available on AI Gateway.  Stream audio in real time with tool calls. …
  • @artificialanlys @artificialanlys on x
    Google has released Gemini 3.8 Live, its new Speech to Speech model, with the Extended Thinking (High) variant debuting at #1 on the Artificial Analysis Speech to Speech Index at 82.6, and #1 on our Tau Voice benchmark implementation at 68.6% Gemini 3.8 Live is @GoogleDeepMind's …
  • @_philschmid Philipp Schmid on x
    Gemini 3.8 Live and 3.8 Live Extended Thinking are here. Following 3.5 Transcribe last month, this continues our focus on real-time voice agents. 🐸🐸 - 82.6 (#1) on Artificial Analysis Quality Index - $0.005/min input and $0.018/min output. - 35.1 (#1) Agentic task completion (τ-b…
  • @thorwebdev @thorwebdev on x
    Say “Hello, ¡Hola!, 你好” to Gemini 3.8 Live & 3.8 Live Extended Thinking. 🎙️⚡ Built for next-gen voice agents that reason, execute, and keep the conversation flowing in real time: ⚙️ Async function calling: Trigger tools in the background without pausing audio streams. 👁️ [video]
  • @google @google on x
    Say “hi
  • @googleai @googleai on x
    Introducing our most advanced Gemini Audio models yet 🗣 Gemini 3.8 Live and 3.8 Live Extended Thinking let you speak, collaborate, and execute tasks seamlessly, meaning conversing with AI just got a lot more natural. So, what's the difference between these two models? Let's break…
  • @cheatyyyy @cheatyyyy on x
    Gemini 3.8 Live is based on Gemini 3 Pro 👀 I thought this would be a flash model since it's latency sensitive, but guess not
  • @valeriawu_ Valeria Wu on x
    We built the 3.8 Live series with enterprise voice agents in mind. Our latest models can combine language switching, visual input, and async function calling for a seamless conversational experience. Build with it on AI Studio today!
  • @officiallogank Logan Kilpatrick on x
    Say hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models, available with frontier price + performance. 3.8 Live supports 97 languages (can seamlessly switch), async tool calls, and more!
  • @newsfromgoogle @newsfromgoogle on x
    Meet our most advanced live dialogue models yet from @GoogleDeepMind: 🔷 Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. 🔷 Gemini 3.8 Live Extended Thinking: Built for high-complexity tasks, with…
  • @googleaistudio @googleaistudio on x
    🔈 today we're introducing two new live dialogue models Gemini 3.8 Live: built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. Gemini 3.8 Live Extended Thinking: built for high-complexity tasks, with increased intellig…
  • @chrisgpt Chris on x
    Wow. Gemini 3.8 Live just dropped right after OpenAI released the GPT-Live-1 API, and Google is already claiming better scores at a significantly lower cost. On this Speech-to-Speech Index, Gemini 3.8 Live Extended Thinking scores 82.6% vs 81.5% for GPT-Live-1 Astra. And it does …
  • Luke Leonhard Luke Leonhard on linkedin
    Congrats to the Gemini Audio team, topping benchmarks (#1 for Tau and #1 for the Artificial Analysis Speech to Speech Index!) with Gemini 3.8 Live models. …
  • @rob-wolfe.com Rob Wolfe on bluesky
    “Live dialogue model.”  —  Don't call them jumped up chat bots or they will get sad [embedded post]
  • r/Bard r on reddit
    Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
  • r/GeminiAI r on reddit
    Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
  • r/singularity r on reddit
    Gemini 3.8 Live & Gemini 3.8 Live Extended Thinking