LeRobot SmolVLA

Robotics / VLAUnavailable
by OtherModel ID: smolvla

HuggingFace's 450M VLA pretrained on 487 community LeRobot datasets. Runs on consumer GPUs.

Status
Unavailable
Input โ†’ output
Text + Image โ†’ Robot actions
Developer
Other
Updated
September 24, 2026

LeRobot SmolVLA is currently unavailable

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

01

Playground

LeRobot SmolVLA

Research model

Currently unavailable

LeRobot SmolVLA is a robotics model (vision-language-action) and cannot be run through the railwail API.

02

About LeRobot SmolVLA

TL;DRAs of September 24, 2026

LeRobot SmolVLA is a model by Other in the Robotics / VLA category. LeRobot SmolVLA is currently not available on Railwail.

Background

About Hugging Face (LeRobot team)

Founded 2016 ยท New York, USA / Paris, France

SmolVLA is the flagship Vision-Language-Action model of Hugging Face's LeRobot project, an open-source robotics framework that brings the Transformers / Datasets philosophy to physical-AI research. SmolVLA was released in mid-2025 as a deliberately compact 450M-parameter VLA designed to be trainable and runnable on consumer hardware while still benefiting from community-scale pretraining. It is trained on 487 publicly contributed LeRobot community datasets - teleoperation episodes uploaded by hobbyists, university labs and small robotics companies - making it the first community-data-driven open VLA. The release includes pretraining and fine-tuning code, model checkpoints under Apache-2.0, and a tightly integrated stack with the LeRobot framework, hf-hub-hosted datasets, and the SO-100 / SO-ARM-100 low-cost robot arms.

Visit Hugging Face (LeRobot team)

Architecture

Compact Vision-Language-Action transformer (action-chunk regression)

SmolVLA is a 450M-parameter transformer that combines a SmolVLM-style vision-language encoder with an action expert that regresses continuous action chunks. The vision-language tower is initialised from the open SmolVLM family (compact VLMs released by Hugging Face) and is responsible for fusing multi-view RGB observations with the natural-language instruction; a smaller action-prediction head consumes the resulting tokens together with proprioception and outputs a short chunk of continuous joint or end-effector actions. The model is pretrained on 487 LeRobot-format community datasets, covering single-arm, dual-arm and mobile-base setups, with a strong tilt toward the popular SO-100 and Koch low-cost teleoperation arms. Post-pretraining, users fine-tune on their own LeRobot recording for a specific robot and task. The whole stack is designed to run pretraining on a few H100s and fine-tuning on a single consumer GPU.

Parameters
450M

Capabilities

  • Compact 450M open VLA pretrained on community data
  • Trained on 487 LeRobot community datasets
  • SmolVLM-style vision-language tower + action expert
  • Continuous action-chunk regression
  • Runs fine-tuning on a single consumer GPU
  • Tight integration with LeRobot framework on Hugging Face
  • Apache-2.0 licence on weights and code
  • Strong baseline for SO-100 and Koch low-cost arms
  • Best for: hobbyists, educators, low-cost robot research.

Training & license

487 publicly contributed LeRobot-format community datasets hosted on the Hugging Face Hub, dominated by teleoperation episodes from low-cost arms (SO-100, Koch) but also including dual-arm and mobile setups. Total scale on the order of millions of frames.

License: Apache-2.0 - fully open weights, code, and datasets (where contributors used compatible licences). Designed for both research and commercial use.

Safety testing: Research / hobbyist artifact - no formal red-teaming. Safety in deployment relies on low-torque hardware (SO-100, etc.) and user-supplied workspace constraints.

Known limitations

  • Modest scale - underperforms 7B VLAs on hard tasks
  • Dataset skew toward SO-100 / Koch low-cost arms
  • Limited language reasoning vs LLM-backed VLAs
  • Sensor coverage is mostly single RGB camera setups
  • Community data quality varies
  • Long-horizon behaviour limited without prompt decomposition
03

Pricing

Currently unavailable. There is no price for this model at the moment, so it cannot be run.

04

API

Call LeRobot SmolVLA with your Railwail API key. Use this model ID in the request:

Not available via the API

Robotics models run on robot hardware, not through the railwail API.

05

Specifications

Model ID
smolvla
Developer
Other
Input
Text, Image
Output
Robot actions
Model size
450M
License
Apache-2.0 - fully open weights, code, and datasets (where contributors used compatible licences). Designed for both research and commercial use.
Catalog entry updated
September 24, 2026

Tags

  • huggingface
  • lerobot
  • vla
  • robotics
  • research-only
  • open-weights
  • small
  • consumer-gpu
06

Use cases

What it is used for

  • Hobbyist robotics with low-cost arms (SO-100, Koch)
  • Educational coursework on VLAs
  • Community-data-driven robot learning research
  • Quick fine-tuning to new tasks on consumer GPUs
  • Reproducible baselines on LeRobot benchmarks
  • On-prem prototypes that need an Apache-2.0 VLA
07

Frequently asked questions

What is LeRobot SmolVLA?

LeRobot SmolVLA is a model by Other in the Robotics / VLA category. It is listed on Railwail but cannot be run at the moment.

How much does LeRobot SmolVLA cost on Railwail?

LeRobot SmolVLA cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

How fast is LeRobot SmolVLA?

There are not enough measured runs of LeRobot SmolVLA on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

When should I use LeRobot SmolVLA?

LeRobot SmolVLA belongs to the Robotics / VLA category. The category page lists the other models of this kind with their prices.

All models in Robotics / VLA

Can LeRobot SmolVLA process images?

Yes. LeRobot SmolVLA accepts images as input in addition to text.

Can I use LeRobot SmolVLA right now?

Currently unavailable. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.