AI Model Timeline

Tracking the accelerating release frequency of frontier AI models.

Gemini Robotics 1.5

Google·

Roboticsroboticscloud

Overview

Vision-language-action model from Google DeepMind that turns camera input and natural-language instructions into robot motor commands. Shipped as a pair: Gemini Robotics 1.5 executes, while Gemini Robotics-ER 1.5 does the embodied reasoning — it can call digital tools such as web search to plan a task before handing execution over. ER 1.5 is available to developers through the Gemini API in Google AI Studio; the VLA itself went to selected partners only.

Capabilities and innovations

Vision-language-action robot controlEmbodied reasoning with tool use (ER 1.5)Motion transfer across robot embodimentsThinks before actingSplit VLA / embodied-reasoning architectureWeb search as a planning step for physical tasks

Benchmarks

No tracked academic benchmark applies to robotics models; the timeline carries no scores and no leadership tracking for them.

  • GPQA Diamond: not reported by the vendor
  • Humanity’s Last Exam (no tools): not reported by the vendor
  • SWE-bench Verified: not reported by the vendor
  • SWE-bench Pro: not reported by the vendor
  • MMLU: not reported by the vendor
  • MMLU-Pro: not reported by the vendor
  • Terminal-Bench 2.x: not reported by the vendor
  • MMMU-Pro: not reported by the vendor

Links

More from Google

Data curated by AI Model Timeline. See the methodology for admission criteria, benchmark eras and source priority.