Gemini Robotics 1.5 introduces AI agents for physical tasks
Google DeepMind launches Gemini Robotics 1.5, a suite of models enabling robots to perceive, plan, and execute complex physical tasks with advanced reasoning capabilities.

What happened?
Google DeepMind has introduced two new models: Gemini Robotics 1.5 and Gemini Robotics-ER 1.5. These models are designed to drive an era of physical agents, allowing robots to use tools, act, and solve complex, multi-step tasks. Gemini Robotics 1.5 functions as a vision-language-action (VLA) model that transforms visual information and instructions into motor commands for robots.
Key facts
| Modellnamn 1 | Gemini Robotics 1.5 |
|---|---|
| Modelltyp 1 | Vision-Language-Action (VLA) |
| Modellnamn 2 | Gemini Robotics-ER 1.5 |
| Modelltyp 2 | Vision-Language Model (VLM) |
| Utvecklare | Google DeepMind |
”We’re powering an era of physical agents — enabling robots to perceive, plan, think, use tools and act to better solve complex, multi-step tasks.”
”Gemini Robotics 1.5 – Our most capable vision-language-action (VLA) model turns visual information and instructions into motor commands for a robot to perform a task.”
”Gemini Robotics-ER 1.5 – Our most capable vision-language model (VLM) reasons about the physical world, natively calls digital tools and creates detailed, multi-step plans to complete a mission.”
Why it matters
The development of Gemini Robotics 1.5 and Gemini Robotics-ER 1.5 represents a step toward more intelligent and versatile robots. These models improve a robot's ability to understand its environment, plan, and execute tasks transparently. This addresses the need for more capable robots that can handle real-world environments and complex scenarios.
Who is affected?
This launch primarily affects developers and researchers in robotics, as well as companies involved in automation and robot development. Users of robotic technology may eventually benefit from more advanced and autonomous systems in fields such as logistics, manufacturing, and service. Google DeepMind is the primary entity behind the development.
Impact on the EU
Not relevant for EU status. The launch is global, and specific EU regulations or particular implications for the EU market are not addressed in the source material.
What else you should know
Gemini Robotics-ER 1.5 is a vision-language model (VLM) that reasons about the physical world, utilizes digital tools, and creates detailed, multi-step plans to complete missions. According to Google DeepMind, this model achieves top-tier results in spatial understanding.
Quick answers about this story
Vad har hänt?
När hände det?
Varför spelar det roll?
Vilka bolag berörs?
The link opens in a new window and leads to the publisher's own site.
Källan har spårats automatiskt från utgivaren via Aheadlines signalkedja.
AI-verktyg i artikeln
Topics
Get similar news straight to your inbox
The reader's room
Send in a question or an addition. The newsroom reads everything before it's published and replies when relevant. No AI-generated text – just people.
Sign in to submit a comment or question.
Read the article through your role
- Decide whether this affects strategy over 6–12 months or is just noise.
- Discuss with leadership: do we own the right question or does ownership need to move?
- Ask: what risk are we taking by NOT acting on this this quarter?
Generated angle — not editorial analysis of "Gemini Robotics 1.5 introduces AI agents for physical tasks"