Octo Small

Robotik / VLAIkke tilgængelig
af OtherModell-ID: octo-small

Compact 27M variant of Octo. Faster inference on consumer GPUs, designed for low-latency control.

Status
Ikke tilgængelig
Input → output
Tekst + Billede → Roboteraktioner
Udvikler
Other
Opdateret
23. september 2026

Octo Small er i øjeblikket utilgængelig

Du kan stadig læse detaljerne på denne side. Vælg en af de tilgængelige alternativer nedenfor for at køre en sammenlignelig model med det samme.

01

Playground

Octo Small

Forskningsmodel

Derzeit nicht verfügbar

Octo Small er en robotik-model (vision-language-action) og kan ikke køres via railwail API.

02

Om Octo Small

Kort sagtFra 23. september 2026

Octo Small er en model af Other i kategorien Robotik / VLA. Octo Small er i øjeblikket ikke tilgængelig på Railwail.

Baggrund

Om UC Berkeley / Stanford (Octo Model Team)

Grundlagt 2023 · Berkeley & Stanford, California, USA

Octo-Small is the compact 27M-parameter variant of the Octo generalist robot policy released by the Octo Model Team - a UC Berkeley + Stanford-led collaboration (Sergey Levine and Chelsea Finn labs) with contributors from Google DeepMind, CMU and Toyota Research Institute. Octo-Small was released alongside Octo-Base in May 2024 to give researchers a CPU/edge-friendly option that still benefits from the same 800k Open-X-Embodiment pretraining recipe. It is widely used in academic teaching, robotics coursework, and rapid prototyping where the 93M Base model is too heavy for the available hardware.

Besøg UC Berkeley / Stanford (Octo Model Team)

Arkitektur

Compact transformer policy with diffusion action head (Vision-Language-Action)

Octo-Small is architecturally identical to Octo-Base but uses a smaller transformer trunk (~27M parameters total). Inputs are tokenised RGB observations from primary and wrist cameras and a T5-base-encoded language instruction, plus learnable readout tokens. The transformer fuses these tokens and emits action latents that are decoded by a diffusion head into continuous 7-DoF end-effector action chunks. It is pretrained on the same ~800k demonstrations from 25 Open-X-Embodiment-compatible datasets across 9 robot embodiments. Despite being ~3.4x smaller than Octo-Base, the small variant retains the diffusion-policy output and the embodiment-agnostic input adapters, making it suitable as a fast baseline and a starting point for fine-tuning to new robots on modest GPUs.

Parametre
27M

Funktioner

  • Compact 27M generalist VLA policy
  • Same training recipe and dataset as Octo-Base
  • Continuous action chunks via diffusion head
  • Runs on a single consumer GPU and many edge devices
  • Fast fine-tuning on new robots / tasks
  • Natural-language instruction conditioning
  • Apache-2.0 open weights and code
  • Reproducible baseline for academic VLA work
  • Best for: edge robotics, teaching, fast iteration.

Træning og licens

~800,000 cross-embodiment robot demonstrations from 25 Open-X-Embodiment datasets (same corpus as Octo-Base). Trained on TPU hardware with the public Octo recipe.

Licens: Apache-2.0 - fully open weights, code, and recipes. Research and commercial use permitted under the licence.

Sikkerhedstests: Research artifact; no formal RSP. Operational safety is the responsibility of downstream deployers (e.g. safety controllers, force limits).

Kendte begrænsninger

  • Lower accuracy than Octo-Base and modern 7B VLAs
  • Limited capacity for long-horizon language reasoning
  • Trained at low image resolution (256x256)
  • Few-shot transfer to drastically new robots still requires demos
  • Single text encoder (T5) limits prompt richness
  • No native multi-modal sensors beyond RGB
03

Priser

Ikke tilgængelig i øjeblikket. Der er i øjeblikket ingen pris for denne model, så den kan ikke køres.

04

API

Kald Octo Small med din Railwail API-nøgle. Brug dette model-ID i anmodningen:

Ikke tilgængelig via API'en

Robotik-modeller køres på robotik-hardware, ikke gennem railwail API'en.

05

Specifikationer

Model-ID
octo-small
Udvikler
Other
Input
Tekst, Billede
Output
Roboteraktioner
Modelstørrelse
27M
Licens
Apache-2.0 - fully open weights, code, and recipes. Research and commercial use permitted under the licence.
Katalogelement opdateret
23. september 2026

Tags

  • berkeley
  • vla
  • robotics
  • research-only
  • open-weights
  • small
  • consumer-gpu
06

Anvendelsestilfælde

Hvad det bruges til

  • Edge / on-robot inference on modest GPUs
  • Coursework and lab-class robot learning
  • Fast prototyping of new manipulation tasks
  • Small-budget academic research
  • Comparison baseline against larger VLAs
  • Starter checkpoint for new-robot fine-tuning
07

Ofte stillede spørgsmål

Hvad er Octo Small?

Octo Small er en model fra Other i kategorien Robotik / VLA. Den er opført på Railwail, men kan ikke køres i øjeblikket.

Hvad koster Octo Small på Railwail?

Octo Small kan ikke køres på Railwail i øjeblikket, så der er ingen aktuel pris. Tilgængelige alternativer med priser er angivet længere nede på denne side.

Hvor hurtig er Octo Small?

Der er endnu ikke nok målte kørsler af Octo Small på Railwail til at angive en udførelsestid. Det afhænger af inputtet, indstillingerne og belastningen hos provideren.

Hvornår skal jeg bruge Octo Small?

Octo Small tilhører kategorien Robotik / VLA. Kategorisiden viser de andre modeller af denne type med deres priser.

Alle modeller: Robotik / VLA

Kan Octo Small behandle billeder?

Ja. Octo Small accepterer billeder som input ud over tekst.

Kan jeg bruge Octo Small lige nu?

Ikke tilgængelig i øjeblikket. Siden forbliver online; tilgængelige alternativer fra samme kategori er angivet længere nede.

Alle modeller via én API

En API-nøgle til alle modeller på Railwail. Forbrug debiteres fra forudbetalte credits, 1 credit = 0,01 US$.