Gemini Robotics-ER

Robotika / VLANem elérhető
Google DeepMind általModell-azonosító: gemini-robotics-er

Embodied-reasoning variant of Gemini Robotics. Enhanced 3D spatial reasoning and trajectory planning.

Állapot
Nem elérhető
Bemenet → kimenet
Szöveg + Kép → Roboterakciók
Fejlesztő
Google DeepMind
Frissítve
2026. szeptember 23.

Gemini Robotics-ER jelenleg nem elérhető

Az oldal részleteit továbbra is elolvashatja. Az alábbi elérhető alternatívák közül válasszon egy hasonló modell azonnali futtatásához.

01

Playground

Gemini Robotics-ER

Kutatási modell

Jelenleg nem elérhető

A Gemini Robotics-ER egy robotikai modell (vision-language-action), és nem futtatható a railwail API-n keresztül.

02

A Gemini Robotics-ER névjegye

Röviden2026. szeptember 23. szerint

A Gemini Robotics-ER a Google DeepMind által fejlesztett modell a Robotika / VLA kategóriában. A Gemini Robotics-ER jelenleg nem érhető el a Railwail-en.

Háttér

A Google DeepMind névjegye

Alapítva: 2010 · London, UK / Mountain View, USA

Google DeepMind announced Gemini Robotics-ER (Embodied Reasoning) alongside Gemini Robotics in March 2025. While Gemini Robotics is the action-producing VLA, Gemini Robotics-ER is the reasoning-focused sibling: a Vision-Language Model variant of Gemini 2.0 specialised for spatial understanding, 3D grounding, point/box prediction, trajectory planning and code generation for robotics. It is designed to be combined with classical motion planners, low-level controllers or with the Gemini Robotics VLA itself. DeepMind positions Gemini Robotics-ER as a 'reasoning brain' that a robot stack can call with multimodal prompts to decompose tasks, locate objects in 2D / 3D, and emit waypoints or Python control code. As with Gemini Robotics, access is limited to research and partner programs.

Google DeepMind meglátogatása

Architektúra

Vision-Language Model for Embodied Reasoning (no end-to-end action head)

Gemini Robotics-ER is a fine-tuned variant of Gemini 2.0 specialised for embodied perception and planning rather than direct control. The architecture preserves the multimodal Transformer backbone of Gemini 2.0 (image, video, text, code) but is post-trained on a curated corpus of embodied tasks: object detection in 2D and 3D, point and bounding-box prediction, grasp prediction, motion-trajectory generation, and code-as-policy outputs that call robot APIs. It can accept egocentric robot camera streams and a natural-language task description, then produce structured outputs such as pixel-space points to grasp, 3D coordinates relative to the camera, planning steps, or Python snippets that drive a downstream controller. In combination with the Gemini Robotics VLA, Robotics-ER provides high-level reasoning while the VLA handles closed-loop low-level actions.

Paraméterek
Undisclosed (Gemini 2.0-class)

Képességek

  • Spatial reasoning over 2D / 3D scenes
  • Point and bounding-box prediction for objects and grasps
  • Trajectory waypoint generation
  • Code-as-policy generation (Python that calls robot APIs)
  • Compositional task planning from natural language
  • Pair with motion planners or with Gemini Robotics VLA
  • Multimodal context: images, video, text, robot state
  • Improved zero-shot performance on embodied QA benchmarks
  • Best for: planning and grounding modules in research robotics stacks.

Képzés és licenc

Gemini 2.0 multimodal pretraining plus embodied post-training on object-detection, 3D grounding, grasp prediction, trajectory planning, and code-generation tasks for robotic control. Draws on Google's internal robot datasets and curated public embodied datasets.

Licenc: Research-only / partner access through Google DeepMind. Not publicly downloadable.

Biztonsági tesztelés: Covered under Google DeepMind's Frontier Safety Framework and Responsible AI evaluations, with specific embodied-AI red-teaming for physical-harm and unsafe-action scenarios.

Ismert korlátozások

  • No direct low-level action output
  • Requires downstream controller or planner
  • Closed model - no public weights or API
  • Spatial reasoning still imperfect on cluttered scenes
  • Latency too high for tight inner control loops
  • Generalisation depends on prompt and tool stack
03

Árak

Jelenleg nem elérhető. Jelenleg nincs ár ehhez a modellhez, ezért nem futtatható.

04

API

Hívja meg a Gemini Robotics-ER modellt a Railwail API-kulcsával. Használja ezt a modell-azonosítót a kérésben:

Nem érhető el az API-n keresztül

A robotika modellek robot hardveren futnak, nem a railwail API-n keresztül.

05

Specifikációk

Modell-azonosító
gemini-robotics-er
Fejlesztő
Google DeepMind
Kategória
Robotika / VLA
Bemenet
Szöveg, Kép
Kimenet
Roboterakciók
Életciklus
Aktuális verzió
Modell mérete
Undisclosed (Gemini 2.0-class)
Licenc
Research-only / partner access through Google DeepMind. Not publicly downloadable.
Katalógus bejegyzés frissítve
2026. szeptember 23.

Címkék

  • google
  • deepmind
  • gemini
  • vla
  • robotics
  • research-only
  • weights-closed
  • embodied-reasoning
06

Felhasználási esetek

Mire használják

  • High-level planner in hierarchical robot stacks
  • Grounding language to 2D / 3D scene primitives
  • Code-as-policy generation for manipulation
  • Embodied QA and instruction parsing
  • Companion to motion planners or VLA controllers
  • Academic embodied-reasoning research
07

Gyakran ismételt kérdések

Mi az a Gemini Robotics-ER?

A Gemini Robotics-ER a Google DeepMind által fejlesztett modell a Robotika / VLA kategóriában. A Railwail katalógusában szerepel, de jelenleg nem futtatható.

Mennyibe kerül a Gemini Robotics-ER a Railwail-on?

A Gemini Robotics-ER jelenleg nem futtatható a Railwail-on, ezért nincs aktuális ár. Az elérhető alternatívák árakkal az oldal alább találhatók.

Milyen gyors a Gemini Robotics-ER?

A Gemini Robotics-ER-nek még nincs elég mért futtatása a Railwail-on ahhoz, hogy futási időt adjunk meg. Ez a bemenettől, a beállításoktól és a szolgáltató terhelésétől függ.

Mikor használjam a Gemini Robotics-ER-t?

A Gemini Robotics-ER a Robotika / VLA kategóriához tartozik. A kategória oldala az ilyen típusú többi modellt listázza az árakkal.

Összes modell: Robotika / VLA

A Gemini Robotics-ER képes képeket feldolgozni?

Igen. A Gemini Robotics-ER szöveg mellett képeket is fogad bemenetként.

Használhatom a Gemini Robotics-ER-t most?

Jelenleg nem elérhető. Az oldal online marad; az ugyanabból a kategóriából elérhető alternatívák az oldal alább találhatók.

Összes modell egy API-n keresztül

Egy API-kulcs a Railwail összes modelljéhez. A használat előre feltöltött kreditek alapján kerül felszámításra, 1 kredit = 0,01 USD.