Google RT-2-X

Robotiikka / VLAEi saatavilla
kehittäjä: Google DeepMindMallin tunnus: rt-2-x

Google's VLA from RT-X collaboration. Trained on Open-X-Embodiment (22 robots, 527 skills), positive transfer.

Tila
Ei saatavilla
Syöte → tulos
Teksti + Kuva → Robottitoiminnot
Kehittäjä
Google DeepMind
Päivitetty
23. syyskuuta 2026

Google RT-2-X ei ole tällä hetkellä saatavilla

Voit silti lukea tiedot tältä sivulta. Valitse yksi alla olevista saatavilla olevista vaihtoehdoista suorittaaksesi vertailukelpoisen mallin heti.

01

Leikkikenttä

Google RT-2-X

Tutkimusmalli

Ei tällä hetkellä saatavilla

Google RT-2-X on robotiikkamalli (vision-language-action) eikä sitä voi suorittaa railwail-ohjelmointirajapinnan kautta.

02

Tietoja: Google RT-2-X

Lyhyesti23. syyskuuta 2026 alkaen

Google RT-2-X on Google DeepMind-kehittäjän malli kategoriasta Robotiikka / VLA. Google RT-2-X ei ole tällä hetkellä saatavilla Railwailissa.

Tausta

Tietoja: Google DeepMind

Perustettu 2010 · London, UK / Mountain View, USA

RT-2-X is Google DeepMind's flagship Robotic Transformer 2 (RT-2) model retrained on the Open-X-Embodiment dataset - the first large-scale, multi-institution effort to assemble a unified robot-learning dataset spanning many labs and robots. Open-X-Embodiment was organised in 2023 by Google DeepMind together with 21+ academic and industry institutions (Stanford, UC Berkeley, CMU, MIT, Toyota Research Institute, etc.), producing the RT-X dataset of ~1 million trajectories across 22 robot embodiments. RT-2-X extends RT-2's Vision-Language-Action recipe - using a PaLM-E / PaLI-X style VLM as backbone and emitting actions as text tokens - to this cross-embodiment corpus, demonstrating positive transfer across robots and a new state of the art on generalist manipulation at the time of release. RT-2-X is research-only and not publicly callable; it remains a key academic reference and the conceptual parent of subsequent open VLAs.

Vieraile sivustolla Google DeepMind

Arkkitehtuuri

Vision-Language-Action transformer (PaLI / PaLM-E backbone, discrete action tokens)

RT-2-X follows the RT-2 design: a large Vision-Language Model (PaLI-X or PaLM-E) is co-fine-tuned on web-scale vision-language data and on robot demonstration data, where robot actions are tokenised as strings of natural-language-like tokens (each action dimension binned and rendered as a token). The same next-token prediction objective therefore trains the model on both internet-scale image-text data and on robot trajectories, allowing the resulting policy to inherit web knowledge (object semantics, OCR, common sense) and route it to motor commands. RT-2-X is the version of this recipe trained on the Open-X-Embodiment / RT-X dataset - ~1 million trajectories across 22 robot embodiments - rather than only on Google's internal kitchen-robot dataset. Public results report 5B and 55B variants, with the 55B model showing the strongest generalisation, especially when prompted with unseen language commands or unseen object combinations.

Parametrit
Up to 55B (RT-2-X variants: 5B and 55B)

Ominaisuudet

  • Generalist VLA trained on Open-X-Embodiment (22 robots)
  • Inherits web-scale knowledge from PaLI / PaLM-E backbones
  • Discrete action-token decoding (text-like vocabulary)
  • Positive transfer across robot embodiments
  • Strong emergent semantic reasoning (e.g. 'pick up the extinct animal')
  • 5B and 55B parameter variants
  • Reference architecture for the modern VLA paradigm
  • Co-training on internet data + robot demos
  • Best for: research, citation, conceptual baseline for VLAs.

Koulutus ja lisenssi

Co-trained on internet-scale vision-language data (PaLI / PaLM-E corpora) plus ~1 million robot trajectories from the Open-X-Embodiment (RT-X) dataset across 22 robot embodiments. Action targets are tokenised continuous controls.

Lisenssi: Research-only - Google DeepMind has not publicly released the RT-2-X weights, code or API. Some Open-X-Embodiment data and smaller RT-X reproductions are available, but the proprietary RT-2-X checkpoints are not.

Turvallisuustestit: Covered under Google's standard Responsible AI process and frontier-safety evaluations. Robot demos were filmed in supervised lab settings with force-limited hardware; no public RSP-style document specific to RT-2-X.

Tunnetut rajoitukset

  • Closed weights - no public API or download
  • Inference latency too high for very fast control loops
  • Discrete tokens limit smoothness vs diffusion / flow-matching policies
  • Cross-embodiment transfer still constrained by action-space differences
  • Long-horizon tasks need external prompt decomposition
  • Dataset skew toward kitchen / tabletop tasks
03

Hinnat

Ei tällä hetkellä saatavilla. Tällä hetkellä tälle mallille ei ole hintaa, joten sitä ei voi suorittaa.

04

API

Kutsu Google RT-2-X Railwail-API-avaimellasi. Käytä tätä mallin tunnusta pyynnössä:

Ei saatavilla API:n kautta

Robotiikan mallit suoritetaan robottilaitteistolla, ei railwail-API:n kautta.

05

Tekniset tiedot

Mallin tunnus
rt-2-x
Kehittäjä
Google DeepMind
Syöte
Teksti, Kuva
Tuloste
Robottitoiminnot
Elinkaari
Ei saatavilla
Mallin koko
Up to 55B (RT-2-X variants: 5B and 55B)
Lisenssi
Research-only - Google DeepMind has not publicly released the RT-2-X weights, code or API. Some Open-X-Embodiment data and smaller RT-X reproductions are available, but the proprietary RT-2-X checkpoints are not.
Luettelokirjaus päivitetty
23. syyskuuta 2026

Tunnisteet

  • google
  • vla
  • robotics
  • research-only
  • weights-closed
06

Käyttötapaukset

Mihin sitä käytetään

  • Reference baseline in VLA / cross-embodiment papers
  • Comparative studies vs OpenVLA / Octo / π-0
  • Academic citation as the canonical large VLA
  • Demonstration of emergent reasoning in robotics
  • Motivating examples for action-token tokenisation research
  • Closed model - no direct deployment use
07

Usein kysytyt kysymykset

Mikä on Google RT-2-X?

Google RT-2-X on Google DeepMindn kehittämä malli Robotiikka / VLA-kategoriassa. Se on listattu Railwailissa, mutta sitä ei voi tällä hetkellä suorittaa.

Paljonko Google RT-2-X maksaa Railwailissa?

Google RT-2-Xta ei voi tällä hetkellä suorittaa Railwailissa, joten nykyistä hintaa ei ole. Saatavilla olevat vaihtoehdot hintojen kanssa on lueteltu tämän sivun alempana.

Kuinka nopea Google RT-2-X on?

Google RT-2-Xlla ei ole vielä tarpeeksi mitattuja suorituksia Railwailissa suoritusajan ilmoittamiseksi. Se riippuu syötteestä, asetuksista ja palveluntarjoajan kuormituksesta.

Milloin Google RT-2-X on oikea valinta?

Google RT-2-X kuuluu Robotiikka / VLA-kategoriaan. Kategorian sivulla on lueteltu muut tämän tyyppisen mallit hintojen kanssa.

Kaikki mallit: Robotiikka / VLA

Voiko Google RT-2-X käsitellä kuvia?

Kyllä. Google RT-2-X hyväksyy kuvia syötteenä tekstin lisäksi.

Voiko Google RT-2-Xa käyttää juuri nyt?

Ei tällä hetkellä saatavilla. Sivu pysyy verkossa; saatavilla olevat vaihtoehdot samasta kategoriasta on lueteltu alempana.

Kaikki mallit yhden API:n kautta

Yksi API-avain kaikille Railwailin malleille. Käyttö laskutetaan prepaid-krediiteistä, 1 krediitti = 0,01 $.