RDT-1B

Robotika / VLANedostupné
od OtherID modelu: rdt-1b

Tsinghua's 1B diffusion-transformer bimanual manipulation policy. Predicts next 64 actions per inference.

Stav
Nedostupné
Vstup → výstup
Text + Obrázok → Akcie robota
Vývojár
Other
Aktualizované
23. septembra 2026

RDT-1B nie je momentálne dostupný

Podrobnosti na tejto stránke si môžete prečítať. Vyberte si jednu z dostupných alternatív nižšie a spustite porovnateľný model hneď.

01

Playground

RDT-1B

Výskumný model

Momentálne nedostupné

RDT-1B je robotický model (vision-language-action) a nie je možné ho spustiť cez railwail API.

02

O RDT-1B

StručneK 23. septembra 2026

RDT-1B je model od Other v kategórii Robotika / VLA. RDT-1B nie je v súčasnosti dostupný na Railwail.

Pozadie

O Tsinghua University (TSAIL / IIIS)

Založené 1911 · Beijing, China

Robotics Diffusion Transformer (RDT) is a generalist bimanual manipulation policy developed at Tsinghua University's TSAIL / Institute for Interdisciplinary Information Sciences (IIIS), home of Jun Zhu's diffusion-modelling group. RDT-1B, introduced in October 2024, is one of the first publicly released billion-scale diffusion-based Vision-Language-Action models, specifically designed for two-arm robots such as Aloha, Mobile Aloha and a custom bimanual platform used by the authors. The project is positioned as a Chinese academic counterpart to π-0 and OpenVLA, with open weights released on Hugging Face under a permissive licence and the explicit aim of enabling fully reproducible bimanual VLA research.

Navštíviť Tsinghua University (TSAIL / IIIS)

Architektúra

Diffusion Transformer Vision-Language-Action policy for bimanual manipulation

RDT-1B is a 1-billion-parameter Diffusion Transformer (DiT) trained as a Vision-Language-Action policy. Inputs are multi-view RGB observations (left + right + overhead), proprioception for both arms and any gripper / mobile-base degrees of freedom, plus a natural-language instruction encoded by a text encoder. The conditioning tokens are fed through a transformer trunk, while a diffusion head denoises continuous action chunks for both arms in a unified action space, allowing dual-arm coordinated motion. Pretraining is done in two stages: a large multi-robot pretraining phase on >1M episodes drawn from public datasets including Open-X-Embodiment and curated bimanual corpora, followed by fine-tuning on the authors' own 6,000-episode bimanual dataset spanning ~300 tasks. RDT-1B reports strong results on dexterous bimanual tasks such as folding T-shirts, pouring, and tool use.

Parametre
1B

Schopnosti

  • 1B-parameter Diffusion Transformer VLA
  • Designed for bimanual manipulation (Aloha-class robots)
  • Trained on >1M cross-embodiment episodes + 6k bimanual demos
  • Continuous action chunks for both arms in a unified space
  • Diffusion head produces smooth coordinated motion
  • Open weights on Hugging Face (permissive licence)
  • Strong results on folding, pouring and tool use
  • Reproducible training and evaluation code
  • Best for: bimanual manipulation research, two-arm fine-tuning.

Tréning a licencia

Pretraining on >1 million robot episodes from Open-X-Embodiment and other public datasets, followed by fine-tuning on a curated bimanual dataset of ~6,000 episodes covering ~300 tasks collected with Aloha-class hardware.

Licencia: Open weights released on Hugging Face under a permissive (CC-BY-NC-style) licence; primarily intended for research use.

Bezpečnostné testy: Academic research artifact; no formal red-teaming or RSP. Safety is the responsibility of downstream deployers (Aloha hardware already includes torque limits and e-stop).

Známe obmedzenia

  • Primarily targets bimanual Aloha-class hardware
  • Requires diffusion sampling at inference (multiple steps)
  • Limited language reasoning compared to LLM-backed VLAs
  • Generalisation to single-arm or mobile platforms needs adapters
  • Mostly indoor-lab evaluation
  • Smaller pretraining text corpus than RT-2-X / OpenVLA
03

Ceny

Momentálne nedostupné. V súčasnosti nie je cena za tento model, preto ho nie je možné spustiť.

04

API

Zavolajte RDT-1B s vaším API kľúčom Railwail. V požiadavke použite toto ID modelu:

Nie je dostupné cez API

Modely robotiky sa spúšťajú na hardvéri robota, nie cez railwail API.

05

Špecifikácie

ID modelu
rdt-1b
Vývojár
Other
Kategória
Robotika / VLA
Vstup
Text, Obrázok
Výstup
Akcie robota
Veľkosť modelu
1B
Licencia
Open weights released on Hugging Face under a permissive (CC-BY-NC-style) licence; primarily intended for research use.
Katalógová položka aktualizovaná
23. septembra 2026

Značky

  • tsinghua
  • vla
  • robotics
  • bimanual
  • research-only
  • open-weights
  • diffusion
06

Prípady použitia

Na čo sa používa

  • Bimanual manipulation research
  • Aloha / Mobile Aloha policy fine-tuning
  • Diffusion-policy ablations and benchmarks
  • Open-source baseline for two-arm VLAs
  • Coordinated dual-arm task learning
  • Academic studies of large diffusion policies
07

Často kladené otázky

Čo je RDT-1B?

RDT-1B je model od Other v kategórii Robotika / VLA. Je uvedený na Railwail, ale momentálne ho nie je možné spustiť.

Koľko stojí RDT-1B na Railwail?

RDT-1B nie je momentálne možné spustiť na Railwail, preto nie je aktuálna cena. Dostupné alternatívy s cenami sú uvedené nižšie na tejto stránke.

Ako rýchly je RDT-1B?

Pre RDT-1B je na Railwail zatiaľ príliš málo meraných spustení na určenie doby spustenia. Závisí to od vstupu, nastavení a zaťaženia u poskytovateľa.

Kedy by som mal používať RDT-1B?

RDT-1B patrí do kategórie Robotika / VLA. Stránka kategórie uvádza ďalšie modely tohto typu s ich cenami.

Všetky modely: Robotika / VLA

Môže RDT-1B spracovávať obrázky?

Áno. RDT-1B akceptuje obrázky ako vstup okrem textu.

Môžem RDT-1B používať práve teraz?

Momentálne nedostupné. Stránka zostáva online; dostupné alternatívy z tej istej kategórie sú uvedené nižšie.

Všetky modely cez jedno API

Jeden API kľúč pre všetky modely na Railwail. Použitie sa účtuje z predplateného kreditu, 1 kredit = 0,01 USD.