Gemini Robotics-ER

Ρομποτική / VLAΜη διαθέσιμο
από Google DeepMindΑναγνωριστικό μοντέλου: gemini-robotics-er

Embodied-reasoning variant of Gemini Robotics. Enhanced 3D spatial reasoning and trajectory planning.

Κατάσταση
Μη διαθέσιμο
Είσοδος → Έξοδος
Κείμενο + Εικόνα → Ενέργειες ρομπότ
Προγραμματιστής
Google DeepMind
Ενημερώθηκε
23 Σεπτεμβρίου 2026

Το Gemini Robotics-ER δεν είναι διαθέσιμο αυτή τη στιγμή

Μπορείτε να διαβάσετε τις λεπτομέρειες σε αυτήν τη σελίδα. Επιλέξτε μία από τις διαθέσιμες εναλλακτικές λύσεις παρακάτω για να εκτελέσετε αμέσως ένα συγκρίσιμο μοντέλο.

01

Playground

Gemini Robotics-ER

Μοντέλο έρευνας

Προς το παρόν μη διαθέσιμο

Το Gemini Robotics-ER είναι ένα μοντέλο ρομποτικής (vision-language-action) και δεν μπορεί να εκτελεστεί μέσω του railwail API.

02

Σχετικά με το Gemini Robotics-ER

ΣύντομαΗμερομηνία: 23 Σεπτεμβρίου 2026

Το Gemini Robotics-ER είναι ένα μοντέλο του Google DeepMind στην κατηγορία Ρομποτική / VLA. Το Gemini Robotics-ER δεν είναι διαθέσιμο στο Railwail αυτή τη στιγμή.

Φόντο

Σχετικά με Google DeepMind

Ιδρύθηκε 2010 · London, UK / Mountain View, USA

Google DeepMind announced Gemini Robotics-ER (Embodied Reasoning) alongside Gemini Robotics in March 2025. While Gemini Robotics is the action-producing VLA, Gemini Robotics-ER is the reasoning-focused sibling: a Vision-Language Model variant of Gemini 2.0 specialised for spatial understanding, 3D grounding, point/box prediction, trajectory planning and code generation for robotics. It is designed to be combined with classical motion planners, low-level controllers or with the Gemini Robotics VLA itself. DeepMind positions Gemini Robotics-ER as a 'reasoning brain' that a robot stack can call with multimodal prompts to decompose tasks, locate objects in 2D / 3D, and emit waypoints or Python control code. As with Gemini Robotics, access is limited to research and partner programs.

Επισκεφθείτε Google DeepMind

Αρχιτεκτονική

Vision-Language Model for Embodied Reasoning (no end-to-end action head)

Gemini Robotics-ER is a fine-tuned variant of Gemini 2.0 specialised for embodied perception and planning rather than direct control. The architecture preserves the multimodal Transformer backbone of Gemini 2.0 (image, video, text, code) but is post-trained on a curated corpus of embodied tasks: object detection in 2D and 3D, point and bounding-box prediction, grasp prediction, motion-trajectory generation, and code-as-policy outputs that call robot APIs. It can accept egocentric robot camera streams and a natural-language task description, then produce structured outputs such as pixel-space points to grasp, 3D coordinates relative to the camera, planning steps, or Python snippets that drive a downstream controller. In combination with the Gemini Robotics VLA, Robotics-ER provides high-level reasoning while the VLA handles closed-loop low-level actions.

Παράμετροι
Undisclosed (Gemini 2.0-class)

Δυνατότητες

  • Spatial reasoning over 2D / 3D scenes
  • Point and bounding-box prediction for objects and grasps
  • Trajectory waypoint generation
  • Code-as-policy generation (Python that calls robot APIs)
  • Compositional task planning from natural language
  • Pair with motion planners or with Gemini Robotics VLA
  • Multimodal context: images, video, text, robot state
  • Improved zero-shot performance on embodied QA benchmarks
  • Best for: planning and grounding modules in research robotics stacks.

Εκπαίδευση & άδεια

Gemini 2.0 multimodal pretraining plus embodied post-training on object-detection, 3D grounding, grasp prediction, trajectory planning, and code-generation tasks for robotic control. Draws on Google's internal robot datasets and curated public embodied datasets.

Άδεια: Research-only / partner access through Google DeepMind. Not publicly downloadable.

Δοκιμές ασφάλειας: Covered under Google DeepMind's Frontier Safety Framework and Responsible AI evaluations, with specific embodied-AI red-teaming for physical-harm and unsafe-action scenarios.

Γνωστοί περιορισμοί

  • No direct low-level action output
  • Requires downstream controller or planner
  • Closed model - no public weights or API
  • Spatial reasoning still imperfect on cluttered scenes
  • Latency too high for tight inner control loops
  • Generalisation depends on prompt and tool stack
03

Τιμολόγηση

Προς το παρόν μη διαθέσιμο. Δεν υπάρχει τιμή για αυτό το μοντέλο αυτή τη στιγμή, επομένως δεν μπορεί να εκτελεστεί.

04

API

Καλέστε το Gemini Robotics-ER με το κλειδί Railwail API. Χρησιμοποιήστε αυτό το ID μοντέλου στο αίτημα:

Δεν είναι διαθέσιμο μέσω του API

Τα μοντέλα ρομποτικής εκτελούνται σε υλικό ρομπότ, όχι μέσω του railwail API.

05

Προδιαγραφές

ID μοντέλου
gemini-robotics-er
Ανάπτυξη
Google DeepMind
Κατηγορία
Ρομποτική / VLA
Είσοδος
Κείμενο, Εικόνα
Έξοδος
Ενέργειες ρομπότ
Κύκλος ζωής
Τρέχουσα έκδοση
Μέγεθος μοντέλου
Undisclosed (Gemini 2.0-class)
Άδεια
Research-only / partner access through Google DeepMind. Not publicly downloadable.
Καταχώρηση καταλόγου ενημερώθηκε
23 Σεπτεμβρίου 2026

Ετικέτες

  • google
  • deepmind
  • gemini
  • vla
  • robotics
  • research-only
  • weights-closed
  • embodied-reasoning
06

Περιπτώσεις χρήσης

Για τι χρησιμοποιείται

  • High-level planner in hierarchical robot stacks
  • Grounding language to 2D / 3D scene primitives
  • Code-as-policy generation for manipulation
  • Embodied QA and instruction parsing
  • Companion to motion planners or VLA controllers
  • Academic embodied-reasoning research
07

Συχνές ερωτήσεις

Τι είναι Gemini Robotics-ER;

Το Gemini Robotics-ER είναι ένα μοντέλο του Google DeepMind στην κατηγορία Ρομποτική / VLA. Είναι καταχωρημένο στο Railwail αλλά δεν μπορεί να εκτελεστεί αυτή τη στιγμή.

Πόσο κοστίζει το Gemini Robotics-ER στο Railwail;

Το Gemini Robotics-ER δεν μπορεί να εκτελεστεί στο Railwail αυτή τη στιγμή, επομένως δεν υπάρχει τρέχουσα τιμή. Διαθέσιμες εναλλακτικές λύσεις με τιμές παρατίθενται παρακάτω σε αυτήν τη σελίδα.

Πόσο γρήγορο είναι το Gemini Robotics-ER;

Δεν υπάρχουν αρκετές μετρημένες εκτελέσεις του Gemini Robotics-ER στο Railwail ακόμα για να δηλωθεί ένας χρόνος εκτέλεσης. Εξαρτάται από την είσοδο, τις ρυθμίσεις και το φορτίο στον πάροχο.

Πότε πρέπει να χρησιμοποιήσω το Gemini Robotics-ER;

Το Gemini Robotics-ER ανήκει στην κατηγορία Ρομποτική / VLA. Η σελίδα κατηγορίας παραθέτει τα άλλα μοντέλα αυτού του είδους με τις τιμές τους.

Όλα τα μοντέλα: Ρομποτική / VLA

Μπορεί το Gemini Robotics-ER να επεξεργαστεί εικόνες;

Ναι. Το Gemini Robotics-ER δέχεται εικόνες ως είσοδο εκτός από κείμενο.

Μπορώ να χρησιμοποιήσω το Gemini Robotics-ER τώρα;

Προς το παρόν μη διαθέσιμο. Η σελίδα παραμένει ενεργή· διαθέσιμες εναλλακτικές λύσεις από την ίδια κατηγορία παρατίθενται παρακάτω.

Όλα τα μοντέλα μέσω ενός API

Ένα κλειδί API για κάθε μοντέλο στο Railwail. Η χρήση χρεώνεται από προπληρωμένα πιστωτικά, 1 πιστωτικό = 0,01 $.