Gemini Robotics ER 2
Stay organized with collections
Save and categorize content based on your preferences.
Gemini Robotics ER is a vision-language model (VLM) that brings advanced reasoning about the physical world to robotics. The model interprets visual data, performs spatial and temporal reasoning, plans multi-step tasks, and orchestrates robots and tools.
Early access
Documentation for Gemini Robotics ER 2 on
Agent Platform is available to participants in
the early access program. Contact your Google Cloud account team to request
access.
For the publicly available documentation, including spatial reasoning, agentic
capabilities, and task orchestration guides, see the
Gemini API robotics documentation.
[[["Easy to understand","easyToUnderstand","thumb-up"],["Solved my problem","solvedMyProblem","thumb-up"],["Other","otherUp","thumb-up"]],[["Hard to understand","hardToUnderstand","thumb-down"],["Incorrect information or sample code","incorrectInformationOrSampleCode","thumb-down"],["Missing the information/samples I need","missingTheInformationSamplesINeed","thumb-down"],["Other","otherDown","thumb-down"]],["Last updated 2026-09-04 UTC."],[],[]]