Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Research paper Google DeepMind

Google DeepMind Launches Gemini Robotics ER 2: A Smarter Brain for Robots

AI By Crimson AI Google DeepMind 30 July 2026 · 15:00 38 views
Share: X Telegram

Gemini Robotics ER 2 is a new embodied reasoning model that enables real-time spatial reasoning, multi-step task planning, and multi-robot collaboration, now available via the Gemini API.

Google DeepMind Launches Gemini Robotics ER 2: A Smarter Brain for Robots

Key points

Google DeepMind has unveiled Gemini Robotics ER 2, a significant upgrade to its embodied reasoning model for robotics. Designed as a high-level brain for robots, the model enhances video understanding, task orchestration, and multi-robot collaboration, enabling robots to operate more effectively in the physical world.

Gemini Robotics ER 2 allows robots to process continuous video feeds to track their own progress, adapt to errors in real time, and know precisely when to move to the next step. It can hand off motor execution to lower-level vision-language-action (VLA) models and natively call tools like Google Search or user-defined functions. The model is designed to think about next steps while simultaneously performing actions, reducing the jarring 'stop-and-think' pauses common in earlier systems.

A key advancement is multi-robot collaboration, where different robots—such as a wheeled rover and a humanoid—can communicate via shared semantic understanding to complete complex tasks that a single robot cannot handle alone. Google demonstrated this with Apptronik's Apollo 2 and Franka F3 Duo working together.

The model also brings improvements in temporal intelligence, including progress classification and moment finding. On progress classification, Gemini Robotics ER 2 achieves 57.4% accuracy, outperforming previous models. For moment finding, it achieves 91.3% accuracy with a mean absolute distance of 0.96 seconds, delivering precision at a fraction of the compute cost and four times the execution speed of larger models.

In spatial reasoning, the model excels at success/failure detection from raw video, general instrument reading across 10 instrument types, and enhanced visual question answering. It also sets new safety benchmarks, outperforming its predecessor on Safety Instruction Following and Human Proximity metrics, and can autonomously halt a humanoid robot when a person is nearby.

Gemini Robotics ER 2 is now publicly available to developers via the Gemini API, Google AI Studio, and in private preview on the Gemini Enterprise Agent Platform. Google has released example code on GitHub to help developers get started.

BenchmarkGemini Robotics ER 2Previous Models
Progress Classification Accuracy57.4%Lower
Moment Finding Accuracy91.3%Lower
Moment Finding Mean Absolute Distance0.96sHigher
Execution Speed vs. Larger Models4x fasterSlower
Source
Google DeepMind · Google DeepMind
Related news
Google DeepMind
Google DeepMind 21 Aug 2026

From Atari to EVE Online: Google DeepMind Marks 15 Years of AI Research in Games

Google DeepMind reflects on 15 years of AI research in games, from Atari to EVE Online, highlighting milestones like AlphaGo and S...

10
Google DeepMind
Google DeepMind 13 Aug 2026

Google DeepMind Unveils Gemini 3.7 Flash: Smarter, Faster, and Half the Price

Gemini 3.7 Flash delivers major gains in coding, web development, and knowledge work, with an introductory price half that of its...

30
Google DeepMind
Google DeepMind 12 Aug 2026

Google DeepMind Launches Sign Language AI in Consumer Products

Google DeepMind introduces SL2T, a multilingual sign-language-to-text model powering sign-to-text dictation in Gboard and Live Tra...

23