LeRobot SmolVLA

Robotica / VLANiet beschikbaar
van OtherModel-ID: smolvla

HuggingFace's 450M VLA pretrained on 487 community LeRobot datasets. Runs on consumer GPUs.

Status
Niet beschikbaar
Invoer → Uitvoer
Tekst + Afbeelding → Robotacties
Ontwikkelaar
Other
Bijgewerkt
24 september 2026

LeRobot SmolVLA is momenteel niet beschikbaar

Je kunt de details op deze pagina nog steeds lezen. Kies een van de beschikbare alternatieven hieronder om direct een vergelijkbaar model uit te voeren.

01

Playground

LeRobot SmolVLA

Onderzoeksmodel

Momenteel niet beschikbaar

LeRobot SmolVLA is een roboticamodel (vision-language-action) en kan niet via de railwail-API worden uitgevoerd.

02

Over LeRobot SmolVLA

SamengevatPer 24 september 2026

LeRobot SmolVLA is een model van Other in de categorie Robotica / VLA. LeRobot SmolVLA is momenteel niet beschikbaar op Railwail.

Achtergrond

Over Hugging Face (LeRobot team)

Opgericht 2016 · New York, USA / Paris, France

SmolVLA is the flagship Vision-Language-Action model of Hugging Face's LeRobot project, an open-source robotics framework that brings the Transformers / Datasets philosophy to physical-AI research. SmolVLA was released in mid-2025 as a deliberately compact 450M-parameter VLA designed to be trainable and runnable on consumer hardware while still benefiting from community-scale pretraining. It is trained on 487 publicly contributed LeRobot community datasets - teleoperation episodes uploaded by hobbyists, university labs and small robotics companies - making it the first community-data-driven open VLA. The release includes pretraining and fine-tuning code, model checkpoints under Apache-2.0, and a tightly integrated stack with the LeRobot framework, hf-hub-hosted datasets, and the SO-100 / SO-ARM-100 low-cost robot arms.

Hugging Face (LeRobot team) bezoeken

Architectuur

Compact Vision-Language-Action transformer (action-chunk regression)

SmolVLA is a 450M-parameter transformer that combines a SmolVLM-style vision-language encoder with an action expert that regresses continuous action chunks. The vision-language tower is initialised from the open SmolVLM family (compact VLMs released by Hugging Face) and is responsible for fusing multi-view RGB observations with the natural-language instruction; a smaller action-prediction head consumes the resulting tokens together with proprioception and outputs a short chunk of continuous joint or end-effector actions. The model is pretrained on 487 LeRobot-format community datasets, covering single-arm, dual-arm and mobile-base setups, with a strong tilt toward the popular SO-100 and Koch low-cost teleoperation arms. Post-pretraining, users fine-tune on their own LeRobot recording for a specific robot and task. The whole stack is designed to run pretraining on a few H100s and fine-tuning on a single consumer GPU.

Parameters
450M

Mogelijkheden

  • Compact 450M open VLA pretrained on community data
  • Trained on 487 LeRobot community datasets
  • SmolVLM-style vision-language tower + action expert
  • Continuous action-chunk regression
  • Runs fine-tuning on a single consumer GPU
  • Tight integration with LeRobot framework on Hugging Face
  • Apache-2.0 licence on weights and code
  • Strong baseline for SO-100 and Koch low-cost arms
  • Best for: hobbyists, educators, low-cost robot research.

Training & licentie

487 publicly contributed LeRobot-format community datasets hosted on the Hugging Face Hub, dominated by teleoperation episodes from low-cost arms (SO-100, Koch) but also including dual-arm and mobile setups. Total scale on the order of millions of frames.

Licentie: Apache-2.0 - fully open weights, code, and datasets (where contributors used compatible licences). Designed for both research and commercial use.

Veiligheidstests: Research / hobbyist artifact - no formal red-teaming. Safety in deployment relies on low-torque hardware (SO-100, etc.) and user-supplied workspace constraints.

Bekende beperkingen

  • Modest scale - underperforms 7B VLAs on hard tasks
  • Dataset skew toward SO-100 / Koch low-cost arms
  • Limited language reasoning vs LLM-backed VLAs
  • Sensor coverage is mostly single RGB camera setups
  • Community data quality varies
  • Long-horizon behaviour limited without prompt decomposition
03

Prijzen

Momenteel niet beschikbaar. Er is momenteel geen prijs voor dit model, dus het kan niet worden uitgevoerd.

04

API

Roep LeRobot SmolVLA aan met je Railwail API-sleutel. Gebruik deze model-ID in het verzoek:

Niet beschikbaar via de API

Robotica-modellen draaien op robotica-hardware, niet via de railwail-API.

05

Specificaties

Model-ID
smolvla
Ontwikkelaar
Other
Invoer
Tekst, Afbeelding
Uitvoer
Robotacties
Modelgrootte
450M
Licentie
Apache-2.0 - fully open weights, code, and datasets (where contributors used compatible licences). Designed for both research and commercial use.
Catalogusitem bijgewerkt
24 september 2026

Tags

  • huggingface
  • lerobot
  • vla
  • robotics
  • research-only
  • open-weights
  • small
  • consumer-gpu
06

Gebruiksscenario's

Waarvoor het wordt gebruikt

  • Hobbyist robotics with low-cost arms (SO-100, Koch)
  • Educational coursework on VLAs
  • Community-data-driven robot learning research
  • Quick fine-tuning to new tasks on consumer GPUs
  • Reproducible baselines on LeRobot benchmarks
  • On-prem prototypes that need an Apache-2.0 VLA
07

Veelgestelde vragen

Wat is LeRobot SmolVLA?

LeRobot SmolVLA is een model van Other in de categorie Robotica / VLA. Het staat in de Railwail-catalogus, maar kan momenteel niet worden uitgevoerd.

Hoeveel kost LeRobot SmolVLA op Railwail?

LeRobot SmolVLA kan momenteel niet op Railwail worden uitgevoerd, dus er is geen huidige prijs. Beschikbare alternatieven met prijzen staan verderop op deze pagina.

Hoe snel is LeRobot SmolVLA?

Er zijn nog niet genoeg gemeten runs van LeRobot SmolVLA op Railwail om een uitvoeringstijd op te geven. Dit hangt af van de invoer, de instellingen en de belasting bij de provider.

Wanneer moet ik LeRobot SmolVLA gebruiken?

LeRobot SmolVLA behoort tot de categorie Robotica / VLA. De categoriepagina bevat de andere modellen van dit type met hun prijzen.

Alle modellen: Robotica / VLA

Kan LeRobot SmolVLA afbeeldingen verwerken?

Ja. LeRobot SmolVLA accepteert afbeeldingen als invoer naast tekst.

Kan ik LeRobot SmolVLA nu gebruiken?

Momenteel niet beschikbaar. De pagina blijft online; beschikbare alternatieven uit dezelfde categorie staan verderop.

Alle modellen via één API

Één API-sleutel voor elk model op Railwail. Gebruik wordt afgerekend via vooraf gekochte credits, 1 credit = US$ 0,01.