Gemini Robotics 2
Our most advanced vision-language-action model (VLA) that converts vision and language input into motor control, enabling a robot to take action
Google DeepMind
Our models enable robots of every shape to think, act, and interact with the world around them. With delicate precision and full-body mastery, they autonomously solve a range of complex tasks – using their intelligence to figure out new situations on the fly.
Slide 1 of 3
Our most advanced vision-language-action model (VLA) that converts vision and language input into motor control, enabling a robot to take action
Our embodied reasoning model: capable of reasoning within physical spaces to make detailed plans, coordinating with humans and other robots
A lightweight version of our VLA model, optimized to run locally on robotic hardware
Our most advanced vision-language-action model (VLA) that converts vision and language input into motor control, enabling a robot to take action
Our embodied reasoning model: capable of reasoning within physical spaces to make detailed plans, coordinating with humans and other robots
A lightweight version of our VLA model, optimized to run locally on robotic hardware
Slide 1 of 6
Most robots are trained to do one specific task over and over. But Gemini Robotics 2 can complete a variety of tasks – even if it hasn’t been trained on them before. It’s able to adapt to new and unfamiliar situations on the fly.
Can be adapted to any bi-arm robot in just a few hours, scaling its intelligence from arms to complex humanoid bodies.
Enabling a new level of dexterity that enables robots to complete delicate actions requiring finesse, like screwing in a light bulb and tying knots.
Controls entire humanoid bodies from feet to fingertips. Enabling robots to perform full-range human-like movements from bending to reaching.
Understands and reasons within the real, physical world. Gemini Robotics 2 pairs deep spatial reasoning with long-horizon planning, enabling robots to map multi-step sequences and complete complex, unfamiliar tasks. Supports multi-robot collaboration, allowing two robots to collaborate, and divide labor to complete a single task.
Understands and responds to everyday commands. Gemini Robotics 2 can explain its approach while performing an action, while users can redirect it without using technical language. This makes it ideal for instructing robots through volatile, hazardous environments.
Slide 1 of 3
Controls entire humanoid robots from feet to fingertips, translating intent into whole-body control to reach, bend, and balance.
Controls complex humanoid hands and parallel grippers to unlock a new level of physical dexterity.
Enables different types of robots to communicate and work together to solve complex workflows a single robot could not do alone.
Become a trusted tester
If you're interested in testing our models, please share a few details to join the waitlist.
We’re helping transform their discoveries into market-leading ventures – and to power an era of intelligent physical AI.