Tuesday · October 6, 2026

 ·  Daily  ·  Newsletters

Resonance network: Quantum InsiderSpace Insider

Google DeepMind Releases Newest Gemini Robotics Reasoning Model Designed to Improve How Robots Interpret and Act in Real World

Credit: Google DeepMind

Insider Brief

  • Google DeepMind released Gemini Robotics-ER 1.6, a robotics reasoning model designed to improve how machines interpret visual inputs, plan tasks and determine task completion in physical environments
  • The model adds capabilities including improved spatial reasoning, multi-view perception and instrument reading, enabling robots to identify objects, understand scenes and interpret gauges in industrial settings
  • Google DeepMind said the system also shows gains in safety and reliability, with improved hazard detection and adherence to physical constraints, and is available via the Gemini API and Google AI Studio

Google DeepMind has released Gemini Robotics-ER 1.6, an updated robotics reasoning model designed to improve how machines interpret and act in physical environments.

According to Google DeepMind, the model provides a high-level reasoning layer for robots, enabling systems to better understand visual inputs, plan tasks and determine when actions are complete. The release reflects continued efforts to connect advances in AI models with real-world robotics use cases, particularly in environments that require spatial awareness and decision-making.

What is Gemini Robotics-ER 1.6?

Gemini Robotics-ER 1.6 is a reasoning-first model built to support embodied AI systems, allowing robots to process visual information and translate it into physical actions. It can also interact with external tools, including search and vision-language-action systems, to support task execution.

Google DeepMind highlighted several areas of improvement over earlier versions of the model:

  • Spatial reasoning and object understanding: Improved ability to identify, count and locate objects, including more accurate detection and fewer errors such as identifying objects that are not present
  • Pointing and relational reasoning: Uses spatial “pointing” as an intermediate step to reason about relationships, trajectories and constraints in a scene
  • Task planning and success detection: Determines whether a task has been completed, allowing robots to decide whether to retry or move to the next step
  • Multi-view perception: Combines inputs from multiple cameras, such as overhead and wrist-mounted views, to build a more complete understanding of dynamic or partially obscured environments
  • Instrument reading: Adds the ability to interpret gauges, thermometers and sight glasses, a capability developed in collaboration with Boston Dynamics for inspection and monitoring tasks

Focus on real-world robotics applications

The company pointed out that the instrument reading capability reflects a practical use case in industrial settings, where robots such as Boston Dynamics’ Spot capture images of equipment that must be interpreted accurately. To do that, the model uses what it calls “agentic vision,” a combination of visual reasoning and intermediate computational steps, such as zooming into images and estimating measurements, to derive readings.

“Capabilities like instrument reading and more reliable task reasoning will enable Spot to see, understand and react to real-world challenges completely autonomously,” Marco da Silva, vice president and general manager of Spot at Boston Dynamics, noted in the announcement.

Improvements in Safety and Reliability

Google DeepMind said the model shows improved adherence to safety constraints, including better identification of potential hazards and more consistent decision-making around what objects can be safely manipulated. The system was also evaluated on tasks involving safety instruction following and risk detection in text and video scenarios.

“On these tasks, our Gemini Robotics-ER models improve over baseline Gemini 3.0 Flash performance (+6% in text, +10% in video) in perceiving injury risks accurately,” the company pointed out.

Availability

Gemini Robotics-ER 1.6 is available through the Gemini API and Google AI Studio, with developer tools and example workflows provided to support integration into robotics systems.

“For robots to be truly helpful in our daily lives and industries, they must do more than follow instructions, they must reason about the physical world,” the company said. “From navigating a complex facility to interpreting the needle on a pressure gauge, a robot’s “embodied reasoning” is what allows it to bridge the gap between digital intelligence and physical action.”

Image credit: Google DeepMind

Greg Bock
About the author
Greg Bock

Greg Bock is an award-winning investigative journalist with more than 25 years of experience in print, digital, and broadcast news. His reporting has spanned crime, politics, business and technology, earning multiple Keystone Awards and a Pennsylvania Association of Broadcasters honors. Through the Associated Press and Nexstar Media Group, his coverage has reached audiences across the United States.

Trending today

Physical AI · Humanoids

Minerva Humanoids Emerges from Stealth with $10M in Funding to Develop Robots for Hazardous Work

Business & Markets

Satlyt Announces $8M Seed to Run AI Models Across Satellites in Orbit

Business & Markets

Augmeta Secures $3M Seed to Assign AI Operators to Every Business KPI

an image of an infinite sign on a blue background
Policy & Government · AI Safety

Meta Disputes Journalist’s Claim That Muse AI Agent Read Private Messages Without Consent

Business & Markets

Hitachi & Agile Robots Partner to Develop Physical AI for Autonomous Manufacturing

The AI economy, every weekday morning

The daily briefing on LinkedIn. Free, one tap to follow.

Exclusives

Exclusive

South Korea’s AI G3 Strategy: Decoded

Scale-ups to Watch

10 Switzerland-Based AI Scale-Ups You Need to Know in 2026

network, blockchain, digital, hand, web, community, artificial, intelligence, steering, interfaces, bokeh, future, digitization, transformation, change, blockchain, blockchain, blockchain, blockchain, blockchain, transformation
Exclusive

Why Crypto Could Be AI’s Payment Layer: BlackRock Sees Stablecoins Connecting Commerce and Compute

AI Predictions
Exclusive

Why AI Predictions Often Get The Technology Right But The Timeline Wrong

Scale-ups to Watch

10 CEE & Baltics-Based AI Scale-Ups You Need to Know in 2026

More in Physical AI

Latest from the same section
Business & Markets

Robo.ai Reports More Than $100M in September Revenue, Forecasts About $600M for 2026

4 hours ago
Physical AI · Humanoids

Minerva Humanoids Emerges from Stealth with $10M in Funding to Develop Robots for Hazardous Work

5 hours ago
Business & Markets

Doosan Robotics Selected for Two South Korean Physical AI Projects With Nearly $74M R&D Budget

6 hours ago
Physical AI

Multiply Labs Raises $75M in Series B Funding to Expand Robotic Drug Manufacturing

6 hours ago

The AI economy, every weekday morning

The daily briefing plus the weekly Scale-ups to watch edition. Free, no spam, unsubscribe any time.