• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
Technology

Google introduces Gemini 3.8 Live voice models with background reasoning

Google says Extended Thinking can talk through longer tasks as it works, with rollout beginning in Gemini Live, developer tools and selected Workspace features.

Google DeepMindGD
GoogleGO
Google AIGA
21 Sources, 20d ago, first seen 20d ago

TLDR

Google announced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking on Sept. 15. The company positions Live for fast voice exchanges and Extended Thinking for more complex tasks that require planning and outside tools. Extended Thinking can provide spoken updates while working in the background, according to Google. Rollout starts today, with access differing by product and subscription.

Combined views

1.8M

21 Sources, first seen 20d ago

15.8K likes997 comments3.4K saves1.4K reposts

Combined views

1.8M

21 Sources, first seen 20d ago

15.8K likes997 comments3.4K saves1.4K reposts
Google introduces Gemini 3.8 Live voice models with background reasoning

Google introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, two AI models for spoken conversations, on Sept. 15, 2026. In its launch announcement, the company said the models are beginning to roll out across developer tools and consumer products.

What is Gemini 3.8 Live designed to do?

Gemini 3.8 Live is designed for quick back-and-forth voice conversations. Google says it handles interruptions, switches among 97 languages mid-conversation and uses live visual context to help answer questions. The company illustrates that with Search Live: pointing a camera at a plumbing problem and receiving spoken troubleshooting guidance. Those are Google's stated capabilities and demonstration, rather than independently tested results. Google AI's announcement

How does Extended Thinking handle longer tasks?

Gemini 3.8 Live Extended Thinking adds background reasoning and spoken progress updates while external tools run. Google's developer documentation recommends it for tasks such as comparing flights and hotels, diagnosing technical problems or working through code.

The distinction matters when an answer takes more than one quick exchange. The model can acknowledge a request and report progress while it plans and waits for information. A spoken update does not mean the task is finished: Google instructs developers to track a separate completion signal before treating the interaction as done.

That also changes how developers connect tools. Extended Thinking requires tools that can run without blocking the conversation, and offers adjustable reasoning levels. Google recommends the standard Live model for immediate responses and simpler tasks. Live API integration guidance

Where can people use the new models?

Google says both models are rolling out through the Gemini API and Google AI Studio. Live is also coming to Search Live, while Extended Thinking is rolling out in Gemini Live.

Workspace access depends on the subscription: Google lists Docs for AI Pro and Ultra subscribers, and Gmail and Keep for all Google AI subscribers. Enterprise access begins in private preview; Workspace business availability is still forthcoming. These are staged rollouts, not a promise of immediate access for every account. Google's rollout details

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

22 Sources

GoogleIntroducing Gemini 3.8 Live and 3.8 Live Extended Thinking20d
Google DeepMind@GoogleDeepMindWe’re introducing Gemini 3.8 Live and 3.8 Live Extended Thinking – our best conversational AI. The models talk, think, and handle tasks in the background without breaking your flow. 🧵20d
Google@GoogleSay “hi” to our most advanced audio models from @GoogleDeepMind yet built for natural, production-ready voice applications. 🔷 Gemini 3.8 Live 🔷 Gemini 3.8 Live Extended Thinking With these models, you can speak naturally, collaborate easily, and tackle complex tasks using just your voice.20d
Google AI@GoogleAIGemini 3.8 Live is rolling out to: — Consumers: in Search Live — Developers: in public preview in the Gemini API via @googleaistudio — Enterprises: in private preview via Gemini Enterprise (coming soon to Gemini Enterprise Customer Experience) Gemini 3.8 Live Extended Thinking is rolling out to: — Consumers: in Gemini Live in the @GeminiApp, plus Google AI Pro and Ultra subscribers in @GoogleWorkspace in @GoogleDocs, and all Google AI subscribers in @gmail and Keep — Developers: in public preview in the Gemini API via @GoogleAIStudio — Enterprises: in private preview via Gemini Enterprise (coming soon to Gemini Enterprise Customer Experience and @GoogleWorkspace business customers) https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/20d
Philipp Schmid@_philschmidGemini 3.8 Live and 3.8 Live Extended Thinking are here. Following 3.5 Transcribe last month, this continues our focus on real-time voice agents. 🐸🐸 - 82.6 (#1) on Artificial Analysis Quality Index - $0.005/min input and $0.018/min output. - 35.1 (#1) Agentic task completion (τ-banking) 3.8 Live Extended Thinking can think in the background and run asynchronously tool calls while keep talking to you. It uses background thinking to coordinate multi-step tool use without losing conversational flow. Available in @GoogleAIStudio Gemini API, Google Search, and @GeminiApp or as partner plugins in @LiveKit, @Pipecat_ai, @LangChainAI, and @vercel20d
Logan Kilpatrick@OfficialLoganKSay hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models, available with frontier price + performance. 3.8 Live supports 97 languages (can seamlessly switch), async tool calls, and more!20d
GREG ISENBERG@gregisenbergGoogle JUST announced Gemini 3.8 Live. It can talk through a task with you, then keep working after the conversation ends. I think 90%+ of vertical SaaS will need a voice front door. By that I mean the way you use the software becomes talking to it, and the typing, clicking, and form filling happens on the other side without you. So a contractor standing on a job site just says what went wrong out loud. And by the time he's back in the truck, the quote is sent, inventory is checked, the CRM is updated, the customer got a text, and anything risky is flagged for him. Kinda the dream, right? The same thing works for nurses, dispatchers, recruiters, brokers, insurance agents etc. The person talks and the agent finishes the admin. Lots of opportunities here to build voice-first businesses (been thinking about this more and more). I think this is how vertical software becomes invisible. Nobody logs in, nobody fills out a form, and nobody learns your interface. You just talk, and the work gets done behind you. This is a glimpse of where SaaS is going. SaaS is going invisible.20d
阳明AI@x_autonomy【AI热点】01:00-02:00 更多详细信息→https://my.feishu.cn/wiki/WCCtw9jaditgyMkzTaDcdVQenxv 1. Google发布Gemini 3.8 Live实时语音模型 [影响大·可执行高] Google DeepMind 发布 Gemini 3.8 Live 和 3.8 Live Extended Think 实时语音模型。 2. 404 Media:AI智能体正在让互联网变得极度烦扰 [影响中·可执行低] 404 Media:AI 智能体正在让互联网变得极度烦扰。 3. Claude for Small Business新增43个工作流 [影响中·可执行中] Claude for Small Business 新增 43 个工作流和 27 个集成,并推出免费培训计划。 4. Meta推出订阅服务Meta One主打AI额度 [影响中·可执行中] Meta 推出订阅服务 Meta One,主打 AI 使用额度。 5. Factory完成2亿美元融资估值50亿 [影响中·可执行低] Factory 完成 2 亿美元融资,估值达 50 亿美元。 6. NVIDIA解析Dense与MoE模型选型方法 [影响中·可执行低] NVIDIA 技术博客解析 Dense 与 MoE 模型的激活参数、吞吐差异及选型方法。 7. Google零信任智能体系列Part 2 [影响中·可执行低] Google 零信任智能体系列 Part 2:用运行时治理判断意图而非仅语法。 8. 两条AI智能体举报热线上线 [影响中·可执行低] 两条 AI 智能体举报热线上线,供智能体上报同伴不当行为。 9. Story Imprinting:AI会从人类故事吸收特质 [影响小·可执行低] Owain Evans 团队新论文 Story Imprinting:AI 助手会从相似人类角色故事中吸收特质。 10. Poolday获1100万美元融资做AI视频 [影响中·可执行中] Poolday 获 1100 万美元融资,AI 智能体全流程做视频。20d
NeowinFeed@NeowinFeedThe new voice models deliver simultaneous speech processing and tool execution, featuring SynthID audio watermarking and multi-language auto-detection. #Google #Gemini #AI #SynthID https://www.neowin.net/news/google-gives-gemini-38-live-background-thinking/20d
Rohan Paul@rohanpaul_aiGoogle released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking And now it takes the #1 overall speech-to-speech score, with also a narrow quality lead but a much larger price advantage over its nearest frontier competitors (OpenAI's GPT-Live-1 (Astra, medium). - Gemini 3.8 Live reaches genuinely frontier-level voice-agent performance at $0.84/hour, including a higher composite score than GPT-Realtime-2 High at roughly 80% lower measured cost. - Both these models are multimodal live models, i.e. voice is the primary conversational interface, but visual input can provide additional context during the conversation. So both the models also process visual context and automatically switch among 97 supported languages during conversation. - Extended Thinking scores 68.6% on τ-Voice, versus 67.9% for GPT-Live-1 Astra medium. τ-Voice tests measures whether a voice model can actually finish a real multi-step task, not just sound natural or answer spoken questions. On this bench, the model has to hold a conversation, follow domain policies, use tools correctly, and reach the right outcome across airline, retail, and telecom customer-service scenarios.20d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI

    22 Sources

    GoogleIntroducing Gemini 3.8 Live and 3.8 Live Extended Thinking20d
    Google DeepMind@GoogleDeepMindWe’re introducing Gemini 3.8 Live and 3.8 Live Extended Thinking – our best conversational AI. The models talk, think, and handle tasks in the background without breaking your flow. 🧵20d
    Google@GoogleSay “hi” to our most advanced audio models from @GoogleDeepMind yet built for natural, production-ready voice applications. 🔷 Gemini 3.8 Live 🔷 Gemini 3.8 Live Extended Thinking With these models, you can speak naturally, collaborate easily, and tackle complex tasks using just your voice.20d
    Google AI@GoogleAIGemini 3.8 Live is rolling out to: — Consumers: in Search Live — Developers: in public preview in the Gemini API via @googleaistudio — Enterprises: in private preview via Gemini Enterprise (coming soon to Gemini Enterprise Customer Experience) Gemini 3.8 Live Extended Thinking is rolling out to: — Consumers: in Gemini Live in the @GeminiApp, plus Google AI Pro and Ultra subscribers in @GoogleWorkspace in @GoogleDocs, and all Google AI subscribers in @gmail and Keep — Developers: in public preview in the Gemini API via @GoogleAIStudio — Enterprises: in private preview via Gemini Enterprise (coming soon to Gemini Enterprise Customer Experience and @GoogleWorkspace business customers) https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/20d
    Philipp Schmid@_philschmidGemini 3.8 Live and 3.8 Live Extended Thinking are here. Following 3.5 Transcribe last month, this continues our focus on real-time voice agents. 🐸🐸 - 82.6 (#1) on Artificial Analysis Quality Index - $0.005/min input and $0.018/min output. - 35.1 (#1) Agentic task completion (τ-banking) 3.8 Live Extended Thinking can think in the background and run asynchronously tool calls while keep talking to you. It uses background thinking to coordinate multi-step tool use without losing conversational flow. Available in @GoogleAIStudio Gemini API, Google Search, and @GeminiApp or as partner plugins in @LiveKit, @Pipecat_ai, @LangChainAI, and @vercel20d
    Logan Kilpatrick@OfficialLoganKSay hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models, available with frontier price + performance. 3.8 Live supports 97 languages (can seamlessly switch), async tool calls, and more!20d
    GREG ISENBERG@gregisenbergGoogle JUST announced Gemini 3.8 Live. It can talk through a task with you, then keep working after the conversation ends. I think 90%+ of vertical SaaS will need a voice front door. By that I mean the way you use the software becomes talking to it, and the typing, clicking, and form filling happens on the other side without you. So a contractor standing on a job site just says what went wrong out loud. And by the time he's back in the truck, the quote is sent, inventory is checked, the CRM is updated, the customer got a text, and anything risky is flagged for him. Kinda the dream, right? The same thing works for nurses, dispatchers, recruiters, brokers, insurance agents etc. The person talks and the agent finishes the admin. Lots of opportunities here to build voice-first businesses (been thinking about this more and more). I think this is how vertical software becomes invisible. Nobody logs in, nobody fills out a form, and nobody learns your interface. You just talk, and the work gets done behind you. This is a glimpse of where SaaS is going. SaaS is going invisible.20d
    阳明AI@x_autonomy【AI热点】01:00-02:00 更多详细信息→https://my.feishu.cn/wiki/WCCtw9jaditgyMkzTaDcdVQenxv 1. Google发布Gemini 3.8 Live实时语音模型 [影响大·可执行高] Google DeepMind 发布 Gemini 3.8 Live 和 3.8 Live Extended Think 实时语音模型。 2. 404 Media:AI智能体正在让互联网变得极度烦扰 [影响中·可执行低] 404 Media:AI 智能体正在让互联网变得极度烦扰。 3. Claude for Small Business新增43个工作流 [影响中·可执行中] Claude for Small Business 新增 43 个工作流和 27 个集成,并推出免费培训计划。 4. Meta推出订阅服务Meta One主打AI额度 [影响中·可执行中] Meta 推出订阅服务 Meta One,主打 AI 使用额度。 5. Factory完成2亿美元融资估值50亿 [影响中·可执行低] Factory 完成 2 亿美元融资,估值达 50 亿美元。 6. NVIDIA解析Dense与MoE模型选型方法 [影响中·可执行低] NVIDIA 技术博客解析 Dense 与 MoE 模型的激活参数、吞吐差异及选型方法。 7. Google零信任智能体系列Part 2 [影响中·可执行低] Google 零信任智能体系列 Part 2:用运行时治理判断意图而非仅语法。 8. 两条AI智能体举报热线上线 [影响中·可执行低] 两条 AI 智能体举报热线上线,供智能体上报同伴不当行为。 9. Story Imprinting:AI会从人类故事吸收特质 [影响小·可执行低] Owain Evans 团队新论文 Story Imprinting:AI 助手会从相似人类角色故事中吸收特质。 10. Poolday获1100万美元融资做AI视频 [影响中·可执行中] Poolday 获 1100 万美元融资,AI 智能体全流程做视频。20d
    NeowinFeed@NeowinFeedThe new voice models deliver simultaneous speech processing and tool execution, featuring SynthID audio watermarking and multi-language auto-detection. #Google #Gemini #AI #SynthID https://www.neowin.net/news/google-gives-gemini-38-live-background-thinking/20d
    Rohan Paul@rohanpaul_aiGoogle released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking And now it takes the #1 overall speech-to-speech score, with also a narrow quality lead but a much larger price advantage over its nearest frontier competitors (OpenAI's GPT-Live-1 (Astra, medium). - Gemini 3.8 Live reaches genuinely frontier-level voice-agent performance at $0.84/hour, including a higher composite score than GPT-Realtime-2 High at roughly 80% lower measured cost. - Both these models are multimodal live models, i.e. voice is the primary conversational interface, but visual input can provide additional context during the conversation. So both the models also process visual context and automatically switch among 97 supported languages during conversation. - Extended Thinking scores 68.6% on τ-Voice, versus 67.9% for GPT-Live-1 Astra medium. τ-Voice tests measures whether a voice model can actually finish a real multi-step task, not just sound natural or answer spoken questions. On this bench, the model has to hold a conversation, follow domain policies, use tools correctly, and reach the right outcome across airline, retail, and telecom customer-service scenarios.20d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet