Announced on July 30, 2026, the new model focuses on improving how robots perceive their surroundings, reason through long-running tasks, coordinate with other robots, and safely interact with people. Think of Gemini Robotics ER 2 as a high-level brain for robots. It allows robots to chat with humans, understand the physical world, and plan multi-step tasks.
Unlike traditional motion-control models, ER 2 acts as a planning and decision-making layer capable of interpreting visual information, natural language instructions, and environment data before delegating execution to robot-specific control systems. This architecture allows the same reasoning engine to work across multiple robot designs without requiring identical hardware.
Key advancement: Multi-robot collaboration. Google DeepMind says Gemini Robotics ER 2 allows different robot types to communicate and coordinate on complex workflows that would otherwise require a single machine to perform every task. The model also continuously processes video to track task progress in real-time and detect when goals are complete.
ER 2 is available now via Google AI Studio and Gemini Enterprise Agent Platform. Google also released Gemini Robotics 2 (full humanoid control) and On-Device 2 (optimized for local execution) alongside ER 2, forming a three-tier robotics stack.