Google Deepmind unveils Gemini Robotics 2 to power robots of all shapes from tabletop arms to humanoids
Google Deepmind ships Gemini Robotics 2—a vision-language-action model that unifies control across tabletop arms to full humanoids.

Why it matters
A new frontier model class: multimodal AI that reasons about and controls physical systems at scale. This is the capability layer that turns robotics from siloed research into a generalizable platform problem.
The key facts
5 to knowGemini Robotics 2 is a vision-language-action model
Designed to control robots across form factors: tabletop arms, mobile bases, full-body humanoids
Includes higher-level reasoning layer for robotics task decomposition
Published July 31, 2026
From Google Deepmind (lab-race relevant: Gemini is Google's frontier model line)
Go to the source
The Decoderthe-decoder.com
Publisher excerpt: Google Deepmind's Gemini Robotics 2 is its most advanced vision-language-action model yet, built to control everything from tabletop robots to full-body humanoids. Gemini Robotics ER 2 adds a higher-level reasoning layer for robotics tasks.