• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Google DeepMind introduces Gemini 3.8 Live and 3.8 Live Extended Thinking

    Google DeepMind says both models automatically detect 97 languages, understand visuals in near real time and use tools in the background without interrupting a chat.

    Google DeepMindGD
    Google AIGA
    Demis HassabisDH
    28 Sources, ,

    TLDR

    Google DeepMind announced the two conversational AI models on September 15, 2026. It says Extended Thinking narrates progress on difficult tasks to keep the conversation going.

    Google said rollout was starting that day, with Gemini 3.8 Live in Search Live and Extended Thinking in Gemini Live, alongside a public preview of both models through the Gemini API in Google AI Studio.

    Artificial Analysis reports that Extended Thinking, tested at High reasoning effort, debuted at No. 1 on its Speech to Speech Index with 82.6 points; the standard model ranked No. 5 with 76.0. It also reports that Extended Thinking at High led its Tau Voice benchmark implementation with 68.6%.

    Combined views

    2.4M

    28 Sources, first seen 22d ago

    Combined views

    2.4M

    28 Sources, first seen 22d ago

    18.7K likes
    22d ago
    first seen 22d ago
    18.7K likes
    1.3K comments
    4.7K saves
    1.8K reposts
    1.3K comments
    4.7K saves
    1.8K reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    28 Sources

    Google DeepMind@GoogleDeepMindWe’re introducing Gemini 3.8 Live and 3.8 Live Extended Thinking – our best conversational AI. The models talk, think, and handle tasks in the background without breaking your flow. 🧵22d
    Google@GoogleSay “hi” to our most advanced audio models from @GoogleDeepMind yet built for natural, production-ready voice applications. 🔷 Gemini 3.8 Live 🔷 Gemini 3.8 Live Extended Thinking With these models, you can speak naturally, collaborate easily, and tackle complex tasks using just your voice.22d
    Google AI@GoogleAIGemini 3.8 Live is rolling out to: — Consumers: in Search Live — Developers: in public preview in the Gemini API via @googleaistudio — Enterprises: in private preview via Gemini Enterprise (coming soon to Gemini Enterprise Customer Experience) Gemini 3.8 Live Extended Thinking is rolling out to: — Consumers: in Gemini Live in the @GeminiApp, plus Google AI Pro and Ultra subscribers in @GoogleWorkspace in @GoogleDocs, and all Google AI subscribers in @gmail and Keep — Developers: in public preview in the Gemini API via @GoogleAIStudio — Enterprises: in private preview via Gemini Enterprise (coming soon to Gemini Enterprise Customer Experience and @GoogleWorkspace business customers) https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/22d
    Philipp Schmid@_philschmidGemini 3.8 Live and 3.8 Live Extended Thinking are here. Following 3.5 Transcribe last month, this continues our focus on real-time voice agents. 🐸🐸 - 82.6 (#1) on Artificial Analysis Quality Index - $0.005/min input and $0.018/min output. - 35.1 (#1) Agentic task completion (τ-banking) 3.8 Live Extended Thinking can think in the background and run asynchronously tool calls while keep talking to you. It uses background thinking to coordinate multi-step tool use without losing conversational flow. Available in @GoogleAIStudio Gemini API, Google Search, and @GeminiApp or as partner plugins in @LiveKit, @Pipecat_ai, @LangChainAI, and @vercel22d
    Logan Kilpatrick@OfficialLoganKSay hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models, available with frontier price + performance. 3.8 Live supports 97 languages (can seamlessly switch), async tool calls, and more!22d
    GREG ISENBERG@gregisenbergGoogle JUST announced Gemini 3.8 Live. It can talk through a task with you, then keep working after the conversation ends. I think 90%+ of vertical SaaS will need a voice front door. By that I mean the way you use the software becomes talking to it, and the typing, clicking, and form filling happens on the other side without you. So a contractor standing on a job site just says what went wrong out loud. And by the time he's back in the truck, the quote is sent, inventory is checked, the CRM is updated, the customer got a text, and anything risky is flagged for him. Kinda the dream, right? The same thing works for nurses, dispatchers, recruiters, brokers, insurance agents etc. The person talks and the agent finishes the admin. Lots of opportunities here to build voice-first businesses (been thinking about this more and more). I think this is how vertical software becomes invisible. Nobody logs in, nobody fills out a form, and nobody learns your interface. You just talk, and the work gets done behind you. This is a glimpse of where SaaS is going. SaaS is going invisible.22d
    Rohan Paul@rohanpaul_aiGoogle released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking And now it takes the #1 overall speech-to-speech score, with also a narrow quality lead but a much larger price advantage over its nearest frontier competitors (OpenAI's GPT-Live-1 (Astra, medium). - Gemini 3.8 Live reaches genuinely frontier-level voice-agent performance at $0.84/hour, including a higher composite score than GPT-Realtime-2 High at roughly 80% lower measured cost. - Both these models are multimodal live models, i.e. voice is the primary conversational interface, but visual input can provide additional context during the conversation. So both the models also process visual context and automatically switch among 97 supported languages during conversation. - Extended Thinking scores 68.6% on τ-Voice, versus 67.9% for GPT-Live-1 Astra medium. τ-Voice tests measures whether a voice model can actually finish a real multi-step task, not just sound natural or answer spoken questions. On this bench, the model has to hold a conversation, follow domain policies, use tools correctly, and reach the right outcome across airline, retail, and telecom customer-service scenarios.22d
    koray kavukcuoglu@koraykvIntroducing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. Voice agents with reasoning capabilities that make conversing with AI feel more intuitive and intelligent. The model can take turns seamlessly, think through complexity, and feel more natural to talk to.22d
    Artificial Analysis@ArtificialAnlysGoogle has released Gemini 3.8 Live, its new Speech to Speech model, with the Extended Thinking (High) variant debuting at #1 on the Artificial Analysis Speech to Speech Index at 82.6, and #1 on our Tau Voice benchmark implementation at 68.6% Gemini 3.8 Live is @GoogleDeepMind's successor to Gemini 3.1 Flash Live, a Speech to Speech model that executes tools and API calls in the background while continuing the conversation. It comes in two variants: the standard Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, which supports configurable reasoning effort. We evaluated the standard model and the Extended Thinking variant at High reasoning effort through the Gemini Live API. Key takeaways: ➤ Speech to Speech Index: Gemini 3.8 Live Extended Thinking (High) debuts at #1 at 82.6, ahead of GPT-Live-1 (Astra, medium) at 81.5, Grok Voice Think Fast 2.0 High at 81.3 and GPT-Live-1 (Sol, low) at 80.1. The standard Gemini 3.8 Live debuts at #5 at 76.0, with both variants up on Gemini 3.1 Flash Live High at 71.5 (+11.1 and +4.5 points). The Index averages Speech Reasoning (Big Bench Audio), Agentic Performance (Tau Voice), Arena Preference and Arena Task Success Rate ➤ Speech Agent Arena: Gemini 3.8 Live ranks #2 in preference at Elo 1083, behind Gemini 3.1 Flash Live (1096) and ahead of GPT-Live-1 (Sol, low) at 1053, and #2 on Task Success Rate at 93.2%, behind Grok Voice Think Fast 2.0 High at 94.6%. Gemini 3.8 Live Extended Thinking (High) trails at Elo 990 with 89.1% task success ➤ Tau Voice: Gemini 3.8 Live Extended Thinking (High) takes the top spot on our Tau Voice benchmark implementation at 68.6%, ahead of GPT-Live-1 (Astra, medium) at 67.9%, GPT-Live-1 (Sol, low) at 59.3% and Grok Voice Think Fast 2.0 High at 56.5% - up from 37.7% for Gemini 3.1 Flash Live High. The standard Gemini 3.8 Live scores 30.1% ➤ Big Bench Audio: Gemini 3.8 Live Extended Thinking (High) scores 97.7% on audio reasoning, ahead of Grok Voice Think Fast 2.0 High at 97.2% and behind Qwen Audio 3.0 Realtime Plus at 99.2%. The standard Gemini 3.8 Live scores 91.7% ➤ Speed: Average Time to First Audio on Big Bench Audio is 1.18 seconds for Gemini 3.8 Live and 1.35 seconds for Extended Thinking (High), both well ahead of Gemini 3.1 Flash Live High (2.99s) and in line with GPT-Live-1 (Sol, low) at 1.24s and GPT-Live-1 (Astra, medium) at 1.34s, though behind Grok Voice Think Fast 2.0 High at 0.70s ➤ Cost: Gemini 3.8 Live costs $0.84 per hour of input audio, the cheapest model in the Index and roughly half the $1.75 of Gemini 3.1 Flash Live High. Extended Thinking (High) costs $3.50 per hour - cheaper than GPT-Live-1 (Sol, low) at $4.47, Grok Voice Think Fast 2.0 High at $4.80 and GPT-Live-1 (Astra, medium) at $5.83, and ~3.1x cheaper than GPT-Realtime-2.1 High at $10.75 See below for more detail ⬇️22d
    Demis Hassabis@demishassabisRT @OfficialLoganK: Say hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models, available with…22d

    28 Sources

    Google DeepMind@GoogleDeepMindWe’re introducing Gemini 3.8 Live and 3.8 Live Extended Thinking – our best conversational AI. The models talk, think, and handle tasks in the background without breaking your flow. 🧵22d
    Google@GoogleSay “hi” to our most advanced audio models from @GoogleDeepMind yet built for natural, production-ready voice applications. 🔷 Gemini 3.8 Live 🔷 Gemini 3.8 Live Extended Thinking With these models, you can speak naturally, collaborate easily, and tackle complex tasks using just your voice.22d
    Google AI@GoogleAIGemini 3.8 Live is rolling out to: — Consumers: in Search Live — Developers: in public preview in the Gemini API via @googleaistudio — Enterprises: in private preview via Gemini Enterprise (coming soon to Gemini Enterprise Customer Experience) Gemini 3.8 Live Extended Thinking is rolling out to: — Consumers: in Gemini Live in the @GeminiApp, plus Google AI Pro and Ultra subscribers in @GoogleWorkspace in @GoogleDocs, and all Google AI subscribers in @gmail and Keep — Developers: in public preview in the Gemini API via @GoogleAIStudio — Enterprises: in private preview via Gemini Enterprise (coming soon to Gemini Enterprise Customer Experience and @GoogleWorkspace business customers) https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/22d
    Philipp Schmid@_philschmidGemini 3.8 Live and 3.8 Live Extended Thinking are here. Following 3.5 Transcribe last month, this continues our focus on real-time voice agents. 🐸🐸 - 82.6 (#1) on Artificial Analysis Quality Index - $0.005/min input and $0.018/min output. - 35.1 (#1) Agentic task completion (τ-banking) 3.8 Live Extended Thinking can think in the background and run asynchronously tool calls while keep talking to you. It uses background thinking to coordinate multi-step tool use without losing conversational flow. Available in @GoogleAIStudio Gemini API, Google Search, and @GeminiApp or as partner plugins in @LiveKit, @Pipecat_ai, @LangChainAI, and @vercel22d
    Logan Kilpatrick@OfficialLoganKSay hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models, available with frontier price + performance. 3.8 Live supports 97 languages (can seamlessly switch), async tool calls, and more!22d
    GREG ISENBERG@gregisenbergGoogle JUST announced Gemini 3.8 Live. It can talk through a task with you, then keep working after the conversation ends. I think 90%+ of vertical SaaS will need a voice front door. By that I mean the way you use the software becomes talking to it, and the typing, clicking, and form filling happens on the other side without you. So a contractor standing on a job site just says what went wrong out loud. And by the time he's back in the truck, the quote is sent, inventory is checked, the CRM is updated, the customer got a text, and anything risky is flagged for him. Kinda the dream, right? The same thing works for nurses, dispatchers, recruiters, brokers, insurance agents etc. The person talks and the agent finishes the admin. Lots of opportunities here to build voice-first businesses (been thinking about this more and more). I think this is how vertical software becomes invisible. Nobody logs in, nobody fills out a form, and nobody learns your interface. You just talk, and the work gets done behind you. This is a glimpse of where SaaS is going. SaaS is going invisible.22d
    Rohan Paul@rohanpaul_aiGoogle released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking And now it takes the #1 overall speech-to-speech score, with also a narrow quality lead but a much larger price advantage over its nearest frontier competitors (OpenAI's GPT-Live-1 (Astra, medium). - Gemini 3.8 Live reaches genuinely frontier-level voice-agent performance at $0.84/hour, including a higher composite score than GPT-Realtime-2 High at roughly 80% lower measured cost. - Both these models are multimodal live models, i.e. voice is the primary conversational interface, but visual input can provide additional context during the conversation. So both the models also process visual context and automatically switch among 97 supported languages during conversation. - Extended Thinking scores 68.6% on τ-Voice, versus 67.9% for GPT-Live-1 Astra medium. τ-Voice tests measures whether a voice model can actually finish a real multi-step task, not just sound natural or answer spoken questions. On this bench, the model has to hold a conversation, follow domain policies, use tools correctly, and reach the right outcome across airline, retail, and telecom customer-service scenarios.22d
    koray kavukcuoglu@koraykvIntroducing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. Voice agents with reasoning capabilities that make conversing with AI feel more intuitive and intelligent. The model can take turns seamlessly, think through complexity, and feel more natural to talk to.22d
    Artificial Analysis@ArtificialAnlysGoogle has released Gemini 3.8 Live, its new Speech to Speech model, with the Extended Thinking (High) variant debuting at #1 on the Artificial Analysis Speech to Speech Index at 82.6, and #1 on our Tau Voice benchmark implementation at 68.6% Gemini 3.8 Live is @GoogleDeepMind's successor to Gemini 3.1 Flash Live, a Speech to Speech model that executes tools and API calls in the background while continuing the conversation. It comes in two variants: the standard Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, which supports configurable reasoning effort. We evaluated the standard model and the Extended Thinking variant at High reasoning effort through the Gemini Live API. Key takeaways: ➤ Speech to Speech Index: Gemini 3.8 Live Extended Thinking (High) debuts at #1 at 82.6, ahead of GPT-Live-1 (Astra, medium) at 81.5, Grok Voice Think Fast 2.0 High at 81.3 and GPT-Live-1 (Sol, low) at 80.1. The standard Gemini 3.8 Live debuts at #5 at 76.0, with both variants up on Gemini 3.1 Flash Live High at 71.5 (+11.1 and +4.5 points). The Index averages Speech Reasoning (Big Bench Audio), Agentic Performance (Tau Voice), Arena Preference and Arena Task Success Rate ➤ Speech Agent Arena: Gemini 3.8 Live ranks #2 in preference at Elo 1083, behind Gemini 3.1 Flash Live (1096) and ahead of GPT-Live-1 (Sol, low) at 1053, and #2 on Task Success Rate at 93.2%, behind Grok Voice Think Fast 2.0 High at 94.6%. Gemini 3.8 Live Extended Thinking (High) trails at Elo 990 with 89.1% task success ➤ Tau Voice: Gemini 3.8 Live Extended Thinking (High) takes the top spot on our Tau Voice benchmark implementation at 68.6%, ahead of GPT-Live-1 (Astra, medium) at 67.9%, GPT-Live-1 (Sol, low) at 59.3% and Grok Voice Think Fast 2.0 High at 56.5% - up from 37.7% for Gemini 3.1 Flash Live High. The standard Gemini 3.8 Live scores 30.1% ➤ Big Bench Audio: Gemini 3.8 Live Extended Thinking (High) scores 97.7% on audio reasoning, ahead of Grok Voice Think Fast 2.0 High at 97.2% and behind Qwen Audio 3.0 Realtime Plus at 99.2%. The standard Gemini 3.8 Live scores 91.7% ➤ Speed: Average Time to First Audio on Big Bench Audio is 1.18 seconds for Gemini 3.8 Live and 1.35 seconds for Extended Thinking (High), both well ahead of Gemini 3.1 Flash Live High (2.99s) and in line with GPT-Live-1 (Sol, low) at 1.24s and GPT-Live-1 (Astra, medium) at 1.34s, though behind Grok Voice Think Fast 2.0 High at 0.70s ➤ Cost: Gemini 3.8 Live costs $0.84 per hour of input audio, the cheapest model in the Index and roughly half the $1.75 of Gemini 3.1 Flash Live High. Extended Thinking (High) costs $3.50 per hour - cheaper than GPT-Live-1 (Sol, low) at $4.47, Grok Voice Think Fast 2.0 High at $4.80 and GPT-Live-1 (Astra, medium) at $5.83, and ~3.1x cheaper than GPT-Realtime-2.1 High at $10.75 See below for more detail ⬇️22d
    Demis Hassabis@demishassabisRT @OfficialLoganK: Say hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models, available with…22d