{"slug": "google-deepmind-says-gemini-robotics-2-enables-full-body-control", "title": "Google DeepMind says Gemini Robotics 2 enables full body control", "summary": "Google DeepMind unveiled Gemini Robotics 2, a vision-language-action model that enables whole-body control, advanced dexterity, and multi-robot collaboration, allowing humanoid robots to perform tasks like cleaning a cluttered room. The updated model can run locally on-device and adapt to new robot bodies in a few hours. DeepMind also released Gemini Robotics ER 2, an embodied reasoning model available on Google AI Studio and in private preview on Gemini Enterprise Agent Platform, and Gemini Robotics On-Device 2 for local execution.", "body_md": "Google DeepMind last week unveiled Gemini Robotics 2, the latest version of its vision-language-action, or VLA, model.\n\nWith the first version, the [company](https://www.therobotreport.com/tag/google-deepmind/) showed how Gemini’s multimodal understanding could drive real-world action. The updated version includes intelligent whole-body control, advanced dexterity, and multi-robot collaboration, DeepMind said.\n\nGemini Robotics 2 enables robots to reason through every movement, unlocking a broad range of tasks. For example, it could enable a [humanoid](https://www.therobotreport.com/category/robots-platforms/humanoids/) to walk, crouch, stretch, and manipulate objects to clean up a cluttered room. The robot could also team up with other robots to finish the job faster.\n\nGemini Robotics 2 can also run locally on-device while adapting to entirely new robotic bodies in just a few hours, claimed DeepMind. In addition to the [VLA](https://deepmind.google/models/gemini-robotics/vla/), the company also released:\n\n[Gemini Robotics ER 2](https://deepmind.google/models/gemini-robotics/embodied-reasoning/): This embodied reasoning (ER) model is a vision language model (VLM) that acts as DeepMind’s agent, enabling robots to communicate with people, understand the physical world, and plan multi-step tasks lasting several minutes. DeepMind is also introducing the ability for robots to work together as a team.[Gemini Robotics On-Device 2](https://deepmind.google/models/gemini-robotics/on-device/): The company optimized this VLA to run locally on robotic devices. The model can now achieve fast adaptation to completely new robot embodiments with a few hours of data, DeepMind said.\n\nGemini Robotics ER 2 is now available on [Google AI Studio](https://aistudio.google.com/prompts/new_chat?model=gemini-robotics-er-2-preview) and in private preview on [Gemini Enterprise Agent Platform](https://accounts.google.com/v3/signin/challenge/pk/presend?TL=AE6FvyYMNK1wiOMUNdZN3K6xdM-ts5lHZ7fUKQPQZJ4meQxHPFxt-rwW9O4b153v&authuser=0&cid=1&continue=https%3A%2F%2Fconsole.cloud.google.com%2Fagent-platform%2Fpublishers%2Fgoogle%2Fmodel-garden%2Fgemini-robotics-er-2-preview-info%3Fpli%3D1&dsh=S-217201330%3A1785534649889644&flowName=GlifWebSignIn&followup=https%3A%2F%2Fconsole.cloud.google.com%2Fagent-platform%2Fpublishers%2Fgoogle%2Fmodel-garden%2Fgemini-robotics-er-2-preview-info%3Fpli%3D1&rart=ANgoxcfKFhZIp5NRU436Lk16G1zdzx1Qb6t1dIgcvPqWXPJw-oN-8zrtJydp_NueYBvM3LhP4aBY6-vCBal104FDrBsmvDOu_O53tURGH_fZ0qkar9tg9k4&service=cloudconsole). The VLA and On-Device models are available to [early-access partners](https://docs.google.com/forms/d/1sM5GqcVMWv-KmKY3TOMpVtQ-lDFeAftQ-d9xQn92jCE/viewform?ts=67cef986&edit_requested=true).\n\n## Gemini Robotics 2 manages the entire robot body\n\nDeepMind’s previous models controlled the humanoid’s upper body to achieve tabletop tasks. Now, Gemini Robotics 2 is expanding physical [AI](https://www.therobotreport.com/category/design-development/ai-cognition/) into whole-body motions.\n\nThe model can now control entire humanoid robots, translating intent into intelligent whole-body control. For example, when controlling [Apptronik](https://therobotreport.com/tag/apptronik)‘s [Apollo 2](https://www.therobotreport.com/apptronik-unveils-apollo-2-flagship-data-collection-training-facility/) humanoid robot, users can ask it to “put the watering can into the green bin in the bottom shelf.”\n\nApollo processes the instruction, walks to the table, and picks up the watering can, takes a few steps to the shelves, and places it precisely in its destination. While DeepMind said robots have more to advance in movement speed, this is an important step towards the skills needed to complete more complex, real-world tasks that require whole-body coordination.\n\nTo be genuinely useful in our homes and workplaces, DeepMind said robots need finesse. Gemini Robotics 2 unlocks a new level of physical dexterity across different [end effectors](https://www.therobotreport.com/category/technologies/grippers-end-effectors/), whether a robot is using hands or grippers.\n\nThe model can now control the five-fingered, 22 degree-of-freedom SharpaWave hand on the Apollo 2 robot to complete delicate actions like tying knots or sealing a ziplock bag. It can also operate standard two-fingered parallel grippers on a [Franka Duo](https://franka.de/fr3-duo) platform to perform complex dexterous tasks such as tight packing.\n\n## DeepMind manages complex tasks that require multiple robots\n\nMost real-world tasks require multiple steps over an extended period of time, said Google DeepMind. To manage this complexity, Gemini Robotics ER 2 serves as the robot’s high-level brain, processing user instructions and communicating with humans.\n\nIt observes the room, reasons about the steps needed to complete the task, coordinates with the VLA to carry out the actions, and tracks progress until the task is done. This setup allows robots to execute complex multi-step tasks, self-correct if a step fails, and generalize to novel situations and goals.\n\nWith the update, DeepMind said it is enabling robots to more reliably execute longer task sequences, lasting several minutes and involving hundreds of decisions. Gemini Robotics ER 2 now understands when tasks begin and end, and it can pinpoint the moment key events occur, marking a step change in progress understanding.\n\nThe company is also introducing multi-robot collaboration. This enables different types of robots to communicate and work together to solve complex workflows a single robot could not do alone.\n\n## Gemini Robotics On-Device 2 targets applications with low connectivity\n\nMany robotic applications need to operate without network latency or internet connectivity. DeepMind built Gemini Robotics On-Device 2 to handle these constraints.\n\nThis model is natively multi-embodiment and inherits DeepMind’s “motion transfer” techniques from [Gemini Robotics 1.5](https://www.therobotreport.com/gemini-robotics-1-5-enables-agentic-experiences-explains-google-deepmind/). The model can now adapt to new bi-arm robot embodiments with just a few hours of adaptation time, typically with less than 200 examples.\n\nThis works even with new embodiments with drastically different shapes, sensors, and degrees of freedom, DeepMind claimed.\n\n## DeepMind reaffirms its commitment to safety\n\nGoogle DeepMind said safety is foundational to its robotics research. Gemini Robotics 2 specifically advances robotics [safety](https://www.therobotreport.com/tag/safety) for navigating the uncertainty of the real world and collaborating alongside humans, the company said.\n\nDeepMind introduced [ASIMOV-Agentic](https://huggingface.co/datasets/google/asimov_agentic/blob/main/README.md), a new benchmark for agentic safety orchestration and uncertainty resolution. For example, it measures the embodied reasoning agent’s ability to refuse unsafe tool calls from a VLA. It also measures the agent’s ability to predict whether a task is possible and to proactively request human intervention when uncertain.\n\nAdditionally, with enhanced embodied reasoning, DeepMind said Gemini Robotics ER 2 is its safest robotics model to date in safety constraint following and human proximity benchmarks. It can better detect when humans are nearby, trigger safety tool calls, and bring the robot to a safe stop if someone approaches too closely. This is a key requirement in [collaborative](https://www.therobotreport.com/category/robots-platforms/collaborative-robot/) safety standards, said DeepMind.", "url": "https://wpnews.pro/news/google-deepmind-says-gemini-robotics-2-enables-full-body-control", "canonical_source": "https://www.therobotreport.com/google-deepmind-says-gemini-robotics-2-enables-full-body-control/", "published_at": "2026-08-02 12:31:32+00:00", "updated_at": "2026-08-02 12:58:01.739489+00:00", "lang": "en", "topics": ["artificial-intelligence", "robotics", "ai-products"], "entities": ["Google DeepMind", "Gemini Robotics 2", "Gemini Robotics ER 2", "Gemini Robotics On-Device 2", "Google AI Studio", "Gemini Enterprise Agent Platform", "Apptronik", "Apollo 2"], "alternates": {"html": "https://wpnews.pro/news/google-deepmind-says-gemini-robotics-2-enables-full-body-control", "markdown": "https://wpnews.pro/news/google-deepmind-says-gemini-robotics-2-enables-full-body-control.md", "text": "https://wpnews.pro/news/google-deepmind-says-gemini-robotics-2-enables-full-body-control.txt", "jsonld": "https://wpnews.pro/news/google-deepmind-says-gemini-robotics-2-enables-full-body-control.jsonld"}}