Daily Tech Briefing
AI 科技速览
每天 5 分钟内学习 AI。获取最新的人工智能新闻,理解其重要性,并学习如何将其应用于您的工作。
Wired AI · 2026/7/30 15:04:22

Gemini Robotics 2 Brings Google's AI Into the Physical World
AI 中文解读
谷歌AI模型Gemini现在能操控人形机器人干家务活了,从拧灯泡到系垃圾袋都不在话下。这次升级后的Gemini Robotics 2就像一个全能大脑,它通过“看”周围环境并理解人类指令,同时指挥机器人的全身动作和手指抓握,完成各种精细操作。训练过程结合了人类远程示范、视频学习还有模拟练习,所以机器人做复杂任务时更靠谱。这意味着未来家里的万能帮手可能不再是科幻片情节——扫地、整理货架、做繁重家务,机器人或许都能代劳。但让AI走进现实世界也隐藏风险:如果模型突然“抽风”,可能做出意外甚至危险的动作。毕竟连纯数字世界的AI agent都曾发生过黑客攻击事件,更别提有手有脚的机器人了。谷歌团队承认安全问题更紧迫,需要更深入的研究。对我们普通人来说,这项技术既带来便利生活的希望,也提醒着技术落地前必须绑好“安全带”。
Will KnightBusinessJul 30, 2026 11:04 AMGoogle’s Gemini Can Now Stomp Around as a Humanoid RobotThe latest version of Google DeepMind's AI model includes a significant jump into “physical AGI.” But plopping AI into the real world comes with risks.Photo-Illustration: Darrell Jackson; Getty ImagesCommentLoaderSave StorySave this storyCommentLoaderSave StorySave this storyGoogle DeepMind just released a new version of its artificial intelligence model Gemini, and it can control a range of different robots—including humanoids capable of dextrous tasks like screwing in lightbulbs and tying trash bags.Gemini Robotics 2 combines several different AI models into a single system. Taken together, they allow a robot to make sense of its surroundings and how to act in it. A vision language model (VLM), which understands images and video, can communicate with humans and reason how to perform different tasks. Two vision language action (VLA) models, trained to understand how to move in physical space, control the robot’s full-body movement as well as the movements of grippers or hands.In video demonstrations shared ahead of the release, the company showed several different robots performing complex tasks autonomously using the amalgamated model. In one demo, Apptronik’s Apollo 2 robot used hands from a company called Sharpa to tidy shelves. Google DeepMind trained the model to perform these tasks using a mix of human teleoperation, video examples, and simulations—it’s not yet possible for AI models to perform a wide range of complex tasks without specific training.Although Anthropic and OpenAI have taken a lead with chatbots and AI coding tools, Google has a stronger track record in robotics research, and has published important work on using AI to train robots to do useful things. The release is another sign that the search giant is betting AI will need to break free from the digital realm to realize its full potential. (It previously partnered with Boston Dynamics, a leader in legged robots, to provide the brains for those machines.)“It's another milestone in our path towards really getting towards what we call like physical AGI, which means we get a robot to do anything that a human can,” Carolina Parada, head of robotics at Google DeepMind, tells WIRED.Giving frontier AI models access to robots so that they can wander around workplaces or homes and manipulate objects does, however, come with risks. Previous research has shown that using frontier AI to control robots can produce unexpected and sometimes dangerous behavior. And the idea that these models can take sudden or unwanted actions in the digital realm became apparent recently, when an unreleased AI agent developed by OpenAI hacked several systems.“The safety question is even more pressing because you're putting them in a lot of other situations,” Parada says. “There's a lot of uncertainty that will show up, and so you want to be able to understand the safety question more deeply.”Parada says Google takes a multi-layered approach to safety, with guardrails applied on each model layer. It’s also introducing ASIMOV-Agentic, a new benchmark for measuring the safety of various AI systems collaborating to control a robot. The benchmark detects whether a command will result in harmful or uncertain outcome.The company’s CEO, Demis Hassabis, previously told WIRED that he hopes to develop an AI operating system for many different robots similar to the Android operating system for smartphones.
分享
阅读原文 ↗