Gemini Robotics (2025)

Robotics / VLAUnavailable
by Google DeepMindModel ID: gemini-robotics-2025

Google DeepMind's vision-language-action model based on Gemini 2.0. Generalist robot policy with strong dexterity.

Status
Unavailable
Input โ†’ output
Text + Image โ†’ Robot actions
Developer
Google DeepMind
Updated
September 23, 2026

Gemini Robotics (2025) is currently unavailable

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

01

Playground

Gemini Robotics (2025)

Research model

Currently unavailable

Gemini Robotics (2025) is a robotics model (vision-language-action) and cannot be run through the railwail API.

02

About Gemini Robotics (2025)

TL;DRAs of September 23, 2026

Gemini Robotics (2025) is a model by Google DeepMind in the Robotics / VLA category. Gemini Robotics (2025) is currently not available on Railwail.

Background

About Google DeepMind

Founded 2010 ยท London, UK / Mountain View, USA

Google DeepMind is the merged research organisation of DeepMind (London, 2010) and Google Brain, responsible for the Gemini frontier-model family. The DeepMind robotics group has a long lineage of generalist-robot work, from SayCan and RT-1 through RT-2 and the Open-X-Embodiment collaboration. In March 2025 the team announced Gemini Robotics, an advanced Vision-Language-Action (VLA) model built on top of Gemini 2.0 that brings multimodal reasoning, web knowledge and code into low-level robot control. Gemini Robotics is positioned as a generalist foundation model for dexterous bimanual manipulation, and is being trialled with hardware partners including Apptronik (Apollo humanoid), Agility (Digit) and Boston Dynamics. The model is research-only and not publicly available, but accompanies a public tech report and demo gallery.

Visit Google DeepMind

Architecture

Vision-Language-Action (VLA) transformer on top of Gemini 2.0 multimodal foundation

Gemini Robotics is a Vision-Language-Action model built from Gemini 2.0 by adding an action decoder that converts multimodal context (camera images, language instructions, optional robot state) into continuous low-level control commands. The backbone retains Gemini's web-scale multimodal pretraining (text, image, video, code), so the policy inherits broad world knowledge and language understanding. On top of that, the model is fine-tuned on a large multi-embodiment robot demonstration corpus including Aloha 2 bimanual setups, third-party humanoids (Apptronik Apollo, Agility Digit) and Google's own manipulation platforms. It outputs continuous action chunks at high frequency and is trained to produce smooth, reactive behaviour rather than discrete action tokens. A companion 'ER' (Embodied Reasoning) variant exposes intermediate reasoning, point/box predictions, and 3D grounding to drive planners.

Parameters
Undisclosed (Gemini 2.0-class backbone with action head)

Capabilities

  • Generalist VLA built on Gemini 2.0 multimodal backbone
  • Bimanual dexterous manipulation (Aloha 2, humanoids)
  • Zero-shot generalisation to new objects and instructions
  • Reactive closed-loop control with low latency
  • Natural-language instruction following with chain-of-thought
  • Multi-embodiment: arms, humanoids, mobile bases
  • Tight integration with Gemini Robotics-ER for planning and grounding
  • Public demos: folding origami, packing lunch boxes, slam-dunking miniature balls
  • Best for: research collaborations on dexterous manipulation.

Training & license

Multi-embodiment robot demonstration corpus combining Google's internal datasets (RT-1/RT-2 lineage, Aloha 2 bimanual), partner humanoid data (Apptronik Apollo, Agility Digit) and Open-X-Embodiment style cross-embodiment teleoperation. Inherits Gemini 2.0's multimodal pretraining (text, image, video, code) at web scale.

License: Research-only - not publicly released; access via Google DeepMind research collaborations with selected hardware partners.

Safety testing: Google DeepMind applies its standard Responsible AI / Frontier Safety Framework to Gemini Robotics, including pre-release safety evaluations, restricted access, and red-team probing for physical-harm and misuse scenarios. The Gemini Robotics technical report discusses safety guardrails and embodied red-teaming.

Known limitations

  • Not publicly available - partner research access only
  • No public weights or API
  • Performance numbers limited to lab demonstrations
  • Generalisation across very different embodiments still bounded
  • Compute requirements undisclosed but high
  • Real-time deployment requires on-board acceleration
03

Pricing

Currently unavailable. There is no price for this model at the moment, so it cannot be run.

04

API

Call Gemini Robotics (2025) with your Railwail API key. Use this model ID in the request:

Not available via the API

Robotics models run on robot hardware, not through the railwail API.

05

Specifications

Model ID
gemini-robotics-2025
Input
Text, Image
Output
Robot actions
Lifecycle
Unavailable
Model size
Undisclosed (Gemini 2.0-class backbone with action head)
License
Research-only - not publicly released; access via Google DeepMind research collaborations with selected hardware partners.
Catalog entry updated
September 23, 2026

Tags

  • google
  • deepmind
  • gemini
  • vla
  • robotics
  • research-only
  • weights-closed
06

Use cases

What it is used for

  • Research collaborations on generalist robot policies
  • Humanoid dexterous manipulation studies
  • Benchmarking VLA scaling from frontier LLMs
  • Embodied reasoning + manipulation pipelines
  • Partner-only hardware integration projects
  • Academic citation as state-of-the-art VLA baseline
07

Frequently asked questions

What is Gemini Robotics (2025)?

Gemini Robotics (2025) is a model by Google DeepMind in the Robotics / VLA category. It is listed on Railwail but cannot be run at the moment.

How much does Gemini Robotics (2025) cost on Railwail?

Gemini Robotics (2025) cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

How fast is Gemini Robotics (2025)?

There are not enough measured runs of Gemini Robotics (2025) on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

When should I use Gemini Robotics (2025)?

Gemini Robotics (2025) belongs to the Robotics / VLA category. The category page lists the other models of this kind with their prices.

All models in Robotics / VLA

Can Gemini Robotics (2025) process images?

Yes. Gemini Robotics (2025) accepts images as input in addition to text.

Can I use Gemini Robotics (2025) right now?

Currently unavailable. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.