觉
AI觉醒星球
Awakening is here
Knowledge File / 全球热点解读
2026-09-16 19 浏览 公开

介绍 Gemini 3.8 Live 与 3.8 Live Extended Thinking

Google DeepMind 发布 Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking,两款实时对话模型在智能与并行推理上升级,支持语音执行复杂任务、实时视觉上下文和后台工具调用,已通过 Gemini API、Google Workspace 和 Gemini 应用开放。

SOURCE / 全球热点解读 MIN / 4 ACCESS / 公开 POST / 2026-09-16 01:05:57

原贴

查看原文
作者:Google DeepMind Blog 来源站点:deepmind.google 原贴时间:

原文

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet. Major upgrades in intelligence and parallel reasoning make them more intuitive to collaborate with and use to execute complex tasks using your voice. Member of Technical Staff, on behalf of the Gemini Audio Team Your browser does not support the audio element. We are launching Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking to make voice interactions more natural, fluid, and intelligent. These models handle complex reasoning, real-time visual context, and background task execution without interrupting your conversation. You can start using these features today through the Gemini API, Google Workspace, and the Gemini app. Check out "Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking" for smarter voice AI. Gemini 3.8 Live offers fast, fluid conversations with real-time visual and language support. Use 3.8 Live Extended Thinking to handle complex tasks while keeping the conversation flowing. These models work in the background to manage tools while you keep chatting. You can try these new features in Google Workspace, Search, and the Gemini app. Google just launched two new AI models that make talking to your devices feel way more natural. They can handle interruptions, switch between languages, and even explain their thought process while they work. Whether you're solving a complex problem or just chatting, the AI now feels like it's actually listening and thinking along with you. It’s a big step toward making AI feel like a real conversation partner. Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent. Gemini 3.8 Live : Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding.

中文翻译

Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking 是我们迄今为止最先进的实时对话模型。

智能和并行推理的重大升级使它们更直观地协作,并用于通过语音执行复杂任务。

技术团队成员,代表 Gemini 音频团队 您的浏览器不支持音频元素。

我们推出 Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking,旨在让语音交互更自然、流畅和智能。

这些模型处理复杂推理、实时视觉上下文和后台任务执行,而不会打断你的对话。

你今天就可以通过 Gemini API、Google Workspace 和 Gemini 应用开始使用这些功能。

查看“介绍 Gemini 3.8 Live 和 3.8 Live Extended Thinking”,了解更智能的语音 AI。

Gemini 3.8 Live 提供快速、流畅的对话,并支持实时视觉和语言。

使用 3.8 Live Extended Thinking 处理复杂任务,同时保持对话流畅。

这些模型在后台工作以管理工具,而你继续聊天。

你可以在 Google Workspace、搜索和 Gemini 应用中尝试这些新功能。

Google 刚刚推出了两款新 AI 模型,让你与设备交谈感觉自然得多。

它们可以处理打断、在语言之间切换,甚至在工作时解释它们的思考过程。

无论你是在解决复杂问题还是只是聊天,AI 现在感觉像是在真正倾听并与你一起思考。

这是让 AI 感觉像真正对话伙伴的一大步。

今天,我们推出两款新模型,它们在接近实时的推理方面带来进步,以更有效地支持语音代理,并让与 AI 对话感觉更直观和智能。

Gemini 3.8 Live:为规模和成本效率而构建,将对话智能与流畅对话和视觉基础相结合。

核心信息

Google DeepMind 发布 Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking,两款实时对话模型在智能与并行推理上升级,支持语音执行复杂任务、实时视觉上下文和后台工具调用,已通过 Gemini API、Google Workspace 和 Gemini 应用开放。

  • Google DeepMind 发布 Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking,两款实时对话模型在智能与并行推理上升级,支持语音执行复杂任务、实时视觉上下文和后台工具调用,已通过 Gemini API、Google Workspace 和 Gemini 应用开放。
  • 原贴提到:Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advan
  • 来源:deepmind.google

详细解读

这是什么信号:Google DeepMind 发布了 Gemini 3.8 Live 与 Gemini 3.8 Live Extended Thinking,明确把“实时语音对话”作为新一轮模型能力的主战场。与单纯提升文本生成不同,这次强调智能与并行推理、实时视觉上下文、后台任务执行,以及对话不中断。这意味着语音 AI 正从“你问我答”升级为可并行处理、可打断、可边聊边办事的交互层。

为什么重要:语音是距离人类最自然的接口,但过去受限于延迟、轮次管理和复杂任务处理。3.8 Live 面向规模与成本效率,Extended Thinking 面向复杂任务,两者组合把“实时对话”和“深度思考”拆成可搭配的能力。这会影响语音代理、客服、车载、教育、无障碍等场景的产品设计。

对谁有价值:开发者可以通过 Gemini API 把语音代理嵌入应用;企业可以在 Google Workspace 中低门槛验证会议、文档和协作场景;产品团队可以重新思考“对话中执行任务”的交互,而不是跳转页面或等待。对做语音入口、智能硬件和多语言服务的团队尤其值得关注。

可以怎么行动:先挑选一个高频语音场景做小范围试点,比如客服分流、现场记录或语音助手;对比 3.8 Live 与 3.8 Live Extended Thinking 在延迟、成本和任务完成度上的差异;设计打断、多语言切换和后台工具调用的测试用例;再决定是否接入 Gemini API、Workspace 或 Gemini 应用。

风险或限制:实时语音模型仍可能产生误识别、幻觉或错误工具调用,尤其在复杂任务中需要人工确认和权限边界;后台执行和实时视觉上下文涉及隐私与合规;多语言和口音覆盖、网络延迟、以及 Extended Thinking 的响应时间都需要实测。原文没有给出具体价格、配额或地区限制,落地前应查看官方文档。

信息差价值

这条内容的真正价值,不只是“有人发布了一个新功能”,而是它揭示了 deepmind.google 背后的产品方向、工作流变化或竞争信号。对 OPC 来说,这种信息可以转化成持续追踪的栏目选题。

如果把《介绍 Gemini 3.8 Live 与 3.8 Live Extended Thinking》放到你的内容系统里,它最大的价值在于帮助读者更快看懂“为什么值得关注”,而不是只看到一条碎片化动态。

参考来源

AI SUMMARY

这篇文章回答了什么

介绍 Gemini 3.8 Live 与 3.8 Live Extended Thinking主要讲什么?

Google DeepMind 发布 Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking,两款实时对话模型在智能与并行推理上升级,支持语音执行复杂任务、实时视觉上下文和后台工具调用,已通过 Gemini API、Google Workspace 和 Gemini 应用开放。

这篇文章最值得关注的要点是什么?

Google DeepMind 发布 Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking,两款实时对话模型在智能与并行推理上升级,支持语音执行复杂任务、实时视觉上下文和后台工具调用,已通…;原贴提到:Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advan;来源:deepmind.google

这篇文章和哪些AI专题相关?

它适合放在AI工具、AI日报、Agent工作流专题里阅读。 关联原因:这篇内容命中「工具、自动化、模型」等主题信号。;这篇内容命中「热点解读」等主题信号。;这篇内容命中「Agent」等主题信号。

阅读这篇文章建议先理解哪些关键词?

建议先理解AI日报、每日AI日报、AI信号、热点解读、BuilderPulse这些关键词,再结合正文判断工具、机会或风险是否值得进入自己的工作流。

上一篇 AIHOT 日报参考 2026-09-16 下一篇 Google DeepMind 发布 Gemini 3.8 Live 和 3.8 Live Extended Thinking