AI觉醒星球
Awakening is here
Knowledge File / 全球热点解读
2026-08-01 5 浏览 公开

Google Deepmind发布Gemini Robotics 2,为从桌面机械臂到人形机器人的各种形态提供动力

Google Deepmind发布Gemini Robotics 2,宣称其最先进的VLA模型,能控制从桌面机械臂到人形机器人,并新增具身推理模型ER 2。

SOURCE / 全球热点解读 MIN / 4 ACCESS / 公开 POST / 2026-08-01 02:25:08

原贴

查看原文
作者:Matthias Bastian 来源站点:the-decoder.com 原贴时间:

原文

Google Deepmind has introduced Gemini Robotics 2, which it calls its most advanced vision-language-action (VLA) model yet. VLA models combine image recognition, language processing, and action control to help robots operate in physical environments. Deepmind says the model can control systems ranging from tabletop arms to full-body humanoid robots. The company describes Gemini Robotics 2 as an "intelligence layer" for a new generation of adaptive robots. It can manage full-body movement, perform fine motor tasks, and coordinate multiple robots, according to Deepmind. Developers can apply for early access through the waitlist . Google Deepmind also introduced Gemini Robotics ER 2 , a model designed for "embodied reasoning." The term refers to understanding the physical world and deciding which actions to take based on that information. ER 2 acts as a higher-level control system for robots and replaces Gemini Robotics ER 1.6, released in April . The new model is available in Google AI Studio . Ad DEC_D_Incontent-1 Ad Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

中文翻译

Google Deepmind推出了Gemini Robotics 2,称其迄今最先进的视觉-语言-动作(VLA)模型。VLA模型结合图像识别、语言处理和动作控制,帮助机器人在物理环境中操作。Deepmind表示,该模型可以控制从桌面机械臂到全身人形机器人的系统。公司称Gemini Robotics 2为新一代自适应机器人的“智能层”。据Deepmind称,它可以管理全身运动、执行精细动作任务,并协调多个机器人。开发者可以通过候补名单申请早期访问。Google Deepmind还推出了Gemini Robotics ER 2,一款专为“具身推理”设计的模型。该术语指理解物理世界并基于该信息决定采取哪些行动。ER 2作为机器人的更高级控制系统,取代了4月发布的Gemini Robotics ER 1.6。该新模型可在Google AI Studio中使用。

核心信息

Google Deepmind发布Gemini Robotics 2,宣称其最先进的VLA模型,能控制从桌面机械臂到人形机器人,并新增具身推理模型ER 2。

  • Google Deepmind发布Gemini Robotics 2,宣称其最先进的VLA模型,能控制从桌面机械臂到人形机器人,并新增具身推理模型ER 2。
  • 原贴提到:Google Deepmind has introduced Gemini Robotics 2, which it calls its mos
  • 来源:the-decoder.com

详细解读

信号:Google Deepmind发布Gemini Robotics 2,标志着VLA模型从实验室走向通用机器人控制平台。作为“智能层”,它试图统一不同形态的机器人控制,从桌面机械臂到人形机器人,这暗示机器人AI正从专用走向通用。

为什么重要:VLA模型将感知、语言和行动结合,使机器人能理解指令并执行物理操作。Gemini Robotics 2宣称支持全身运动和精细操作,还具备多机器人协调能力。同时发布的ER 2强化了“具身推理”,让机器人不仅能执行,还能基于物理世界信息做决策。这可能加速机器人在制造、物流、服务等领域的落地。

对谁有价值:机器人开发商可借助该模型快速构建适配各种硬件的控制层;AI研究者可探索多模态大模型与机器人结合的新范式;企业用户可评估未来采用此类技术优化自动化流程的潜力。

可以怎么行动:开发者可申请waitlist获取早期访问,并在Google AI Studio中测试ER 2模型。建议先从小型机械臂场景试水,验证模型在特定任务的性能。同时关注文档和社区评测,积累实际使用经验。

风险或限制:当前处于早期访问阶段,模型能力与鲁棒性尚未经过大规模验证。控制不同形态机器人可能仍需要适配工作,且具身推理的决策可靠性不足。此外,多机器人协调与物理交互可能带来安全风险,需要严格测试。

信息差价值

这条内容的真正价值,不只是“有人发布了一个新功能”,而是它揭示了 the-decoder.com 背后的产品方向、工作流变化或竞争信号。对 OPC 来说,这种信息可以转化成持续追踪的栏目选题。

如果把《Google Deepmind发布Gemini Robotics 2,为从桌面机械臂到人形机器人的各种形态提供动力》放到你的内容系统里,它最大的价值在于帮助读者更快看懂“为什么值得关注”,而不是只看到一条碎片化动态。

参考来源

上一篇 Tailscale 未能阻止 Hugging Face 入侵事件复盘 下一篇 npm 限制绕过2FA的粒度访问令牌