Google DeepMind just dropped what might be the most ambitious robotics announcement of the year. Gemini Robotics 2, unveiled on July 30, is designed to be a single AI brain that can operate inside virtually any robot body, from humanoids to industrial arms, adapting to new hardware in hours rather than months.

What Gemini Robotics 2 actually does

The system is built around three distinct models, each handling a different layer of robot intelligence. The vision-language-action (VLA) model manages whole-body control and fine motor skills. The embodied reasoning (ER) model acts as the high-level brain, orchestrating multi-step task sequences that can stretch over several minutes. And an on-device variant runs locally, meaning the robot doesn’t need to phone home to a cloud server every time it picks up a cup.

The system is hardware-agnostic. A single model checkpoint can operate completely different robotic platforms. Google demonstrated this across Apptronik’s Apollo 2 humanoid and Franka’s robotic arm systems, running the same intelligence layer on radically different physical forms.

Gemini Robotics 2 can learn to operate a new robot embodiment with fewer than 200 training examples, achievable in a few hours of data collection. For context, traditional robotics AI systems often require thousands or tens of thousands of demonstrations to achieve basic competence on a single platform.