RDT-1B

Robotika / VLANedostupné
od OtherID modelu: rdt-1b

Tsinghua's 1B diffusion-transformer bimanual manipulation policy. Predicts next 64 actions per inference.

Stav
Nedostupné
Vstup → výstup
Text + Obrázek → Akce robota
Vývojář
Other
Aktualizováno
23. září 2026

RDT-1B není momentálně dostupný

Podrobnosti na této stránce si můžete přečíst. Vyberte jednu z dostupných alternativ níže a spusťte porovnatelný model hned.

01

Playground

RDT-1B

Výzkumný model

Momentálně nedostupné

RDT-1B je robotický model (vision-language-action) a nelze jej spustit přes railwail API.

02

O RDT-1B

StručněStav: 23. září 2026

RDT-1B je model od Other v kategorii Robotika / VLA. RDT-1B není na Railwail v současné době dostupný.

Pozadí

O Tsinghua University (TSAIL / IIIS)

Založeno 1911 · Beijing, China

Robotics Diffusion Transformer (RDT) is a generalist bimanual manipulation policy developed at Tsinghua University's TSAIL / Institute for Interdisciplinary Information Sciences (IIIS), home of Jun Zhu's diffusion-modelling group. RDT-1B, introduced in October 2024, is one of the first publicly released billion-scale diffusion-based Vision-Language-Action models, specifically designed for two-arm robots such as Aloha, Mobile Aloha and a custom bimanual platform used by the authors. The project is positioned as a Chinese academic counterpart to π-0 and OpenVLA, with open weights released on Hugging Face under a permissive licence and the explicit aim of enabling fully reproducible bimanual VLA research.

Navštívit Tsinghua University (TSAIL / IIIS)

Architektura

Diffusion Transformer Vision-Language-Action policy for bimanual manipulation

RDT-1B is a 1-billion-parameter Diffusion Transformer (DiT) trained as a Vision-Language-Action policy. Inputs are multi-view RGB observations (left + right + overhead), proprioception for both arms and any gripper / mobile-base degrees of freedom, plus a natural-language instruction encoded by a text encoder. The conditioning tokens are fed through a transformer trunk, while a diffusion head denoises continuous action chunks for both arms in a unified action space, allowing dual-arm coordinated motion. Pretraining is done in two stages: a large multi-robot pretraining phase on >1M episodes drawn from public datasets including Open-X-Embodiment and curated bimanual corpora, followed by fine-tuning on the authors' own 6,000-episode bimanual dataset spanning ~300 tasks. RDT-1B reports strong results on dexterous bimanual tasks such as folding T-shirts, pouring, and tool use.

Parametry
1B

Schopnosti

  • 1B-parameter Diffusion Transformer VLA
  • Designed for bimanual manipulation (Aloha-class robots)
  • Trained on >1M cross-embodiment episodes + 6k bimanual demos
  • Continuous action chunks for both arms in a unified space
  • Diffusion head produces smooth coordinated motion
  • Open weights on Hugging Face (permissive licence)
  • Strong results on folding, pouring and tool use
  • Reproducible training and evaluation code
  • Best for: bimanual manipulation research, two-arm fine-tuning.

Trénování a licence

Pretraining on >1 million robot episodes from Open-X-Embodiment and other public datasets, followed by fine-tuning on a curated bimanual dataset of ~6,000 episodes covering ~300 tasks collected with Aloha-class hardware.

Licence: Open weights released on Hugging Face under a permissive (CC-BY-NC-style) licence; primarily intended for research use.

Bezpečnostní testy: Academic research artifact; no formal red-teaming or RSP. Safety is the responsibility of downstream deployers (Aloha hardware already includes torque limits and e-stop).

Známá omezení

  • Primarily targets bimanual Aloha-class hardware
  • Requires diffusion sampling at inference (multiple steps)
  • Limited language reasoning compared to LLM-backed VLAs
  • Generalisation to single-arm or mobile platforms needs adapters
  • Mostly indoor-lab evaluation
  • Smaller pretraining text corpus than RT-2-X / OpenVLA
03

Ceny

Momentálně nedostupné. Pro tento model momentálně není cena, takže jej nelze spustit.

04

API

Volejte RDT-1B s vaším API klíčem Railwail. V požadavku použijte toto ID modelu:

Není dostupné přes API

Robotické modely běží na robotickém hardwaru, ne přes railwail API.

05

Specifikace

ID modelu
rdt-1b
Vývojář
Other
Vstup
Text, Obrázek
Výstup
Akce robota
Velikost modelu
1B
Licence
Open weights released on Hugging Face under a permissive (CC-BY-NC-style) licence; primarily intended for research use.
Záznam v katalogu aktualizován
23. září 2026

Štítky

  • tsinghua
  • vla
  • robotics
  • bimanual
  • research-only
  • open-weights
  • diffusion
06

Případy použití

K čemu se používá

  • Bimanual manipulation research
  • Aloha / Mobile Aloha policy fine-tuning
  • Diffusion-policy ablations and benchmarks
  • Open-source baseline for two-arm VLAs
  • Coordinated dual-arm task learning
  • Academic studies of large diffusion policies
07

Často kladené otázky

Co je RDT-1B?

RDT-1B je model od Other v kategorii Robotika / VLA. Je uveden na Railwail, ale momentálně jej nelze spustit.

Kolik stojí RDT-1B na Railwail?

RDT-1B se momentálně na Railwail nedá spustit, takže není aktuální cena. Dostupné alternativy s cenami jsou uvedeny dále na této stránce.

Jak rychlý je RDT-1B?

Pro RDT-1B je na Railwail zatím příliš málo naměřených spuštění na uvedení doby běhu. Závisí na vstupu, nastavení a zátěži u poskytovatele.

Kdy mám použít RDT-1B?

RDT-1B patří do kategorie Robotika / VLA. Stránka kategorie uvádí další modely tohoto typu s jejich cenami.

Všechny modely: Robotika / VLA

Může RDT-1B zpracovávat obrázky?

Ano. RDT-1B přijímá obrázky jako vstup kromě textu.

Mohu RDT-1B používat hned teď?

Momentálně nedostupné. Stránka zůstává online; dostupné alternativy ze stejné kategorie jsou uvedeny dále.

Všechny modely přes jednu API

Jeden API klíč pro všechny modely na Railwail. Použití se účtuje z předplacených kreditů, 1 kredit = 0,01 US$.