Gemini Robotics 2 is presented as a generalist robotics model that translates natural-language goals and visual understanding into coordinated physical action. The demonstrations emphasize whole-body balance, dexterous manipulation and adaptation rather than a robot repeating one narrowly programmed movement.
The release focuses on three capabilities. Whole-body control coordinates movement from feet to fingertips, dexterous hands handle tasks such as closing a bag or unscrewing a bulb, and separate robots use their own model copies to reason about when and how to help each other.
A high-level embodied reasoning model interprets the scene, instructions and task state, then calls a vision-language-action model to produce movement. The examples show robots dividing a garage-cleanup task, handing work from a humanoid to another robot, packing sports equipment and retrying after a failed action.
The demonstrations suggest a path toward robots that can operate in spaces designed for human bodies and respond to changing conditions. They remain controlled demonstrations supplied by the developer, so they show intended capabilities rather than independent evidence of reliability across unrestricted real-world settings.
Watch on YouTube



