Octo Small

Robotikk / VLAIkke tilgjengelig
av OtherModell-ID: octo-small

Compact 27M variant of Octo. Faster inference on consumer GPUs, designed for low-latency control.

Status
Ikke tilgjengelig
Input → output
Tekst + Bilde → Roboteraksjoner
Utvikler
Other
Oppdatert
23. september 2026

Octo Small er for øyeblikket utilgjengelig

Du kan fortsatt lese detaljene på denne siden. Velg ett av de tilgjengelige alternativene nedenfor for å kjøre en sammenlignbar modell med en gang.

01

Playground

Octo Small

Forskningsmodell

Ikke tilgjengelig for øyeblikket

Octo Small er en robotikkmodell (vision-language-action) og kan ikke kjøres gjennom railwail API.

02

Om Octo Small

Kort sagtPer 23. september 2026

Octo Small er en modell fra Other i kategorien Robotikk / VLA. Octo Small er for øyeblikket ikke tilgjengelig på Railwail.

Bakgrunn

Om UC Berkeley / Stanford (Octo Model Team)

Grunnlagt 2023 · Berkeley & Stanford, California, USA

Octo-Small is the compact 27M-parameter variant of the Octo generalist robot policy released by the Octo Model Team - a UC Berkeley + Stanford-led collaboration (Sergey Levine and Chelsea Finn labs) with contributors from Google DeepMind, CMU and Toyota Research Institute. Octo-Small was released alongside Octo-Base in May 2024 to give researchers a CPU/edge-friendly option that still benefits from the same 800k Open-X-Embodiment pretraining recipe. It is widely used in academic teaching, robotics coursework, and rapid prototyping where the 93M Base model is too heavy for the available hardware.

Besøk UC Berkeley / Stanford (Octo Model Team)

Arkitektur

Compact transformer policy with diffusion action head (Vision-Language-Action)

Octo-Small is architecturally identical to Octo-Base but uses a smaller transformer trunk (~27M parameters total). Inputs are tokenised RGB observations from primary and wrist cameras and a T5-base-encoded language instruction, plus learnable readout tokens. The transformer fuses these tokens and emits action latents that are decoded by a diffusion head into continuous 7-DoF end-effector action chunks. It is pretrained on the same ~800k demonstrations from 25 Open-X-Embodiment-compatible datasets across 9 robot embodiments. Despite being ~3.4x smaller than Octo-Base, the small variant retains the diffusion-policy output and the embodiment-agnostic input adapters, making it suitable as a fast baseline and a starting point for fine-tuning to new robots on modest GPUs.

Parametere
27M

Evner

  • Compact 27M generalist VLA policy
  • Same training recipe and dataset as Octo-Base
  • Continuous action chunks via diffusion head
  • Runs on a single consumer GPU and many edge devices
  • Fast fine-tuning on new robots / tasks
  • Natural-language instruction conditioning
  • Apache-2.0 open weights and code
  • Reproducible baseline for academic VLA work
  • Best for: edge robotics, teaching, fast iteration.

Trening og lisens

~800,000 cross-embodiment robot demonstrations from 25 Open-X-Embodiment datasets (same corpus as Octo-Base). Trained on TPU hardware with the public Octo recipe.

Lisens: Apache-2.0 - fully open weights, code, and recipes. Research and commercial use permitted under the licence.

Sikkerhetstesting: Research artifact; no formal RSP. Operational safety is the responsibility of downstream deployers (e.g. safety controllers, force limits).

Kjente begrensninger

  • Lower accuracy than Octo-Base and modern 7B VLAs
  • Limited capacity for long-horizon language reasoning
  • Trained at low image resolution (256x256)
  • Few-shot transfer to drastically new robots still requires demos
  • Single text encoder (T5) limits prompt richness
  • No native multi-modal sensors beyond RGB
03

Priser

Ikke tilgjengelig for øyeblikket. Det er ingen pris for denne modellen for øyeblikket, så den kan ikke kjøres.

04

API

Ring Octo Small med din Railwail API-nøkkel. Bruk denne modell-IDen i forespørselen:

Ikke tilgjengelig via API-et

Robotikk-modeller kjører på robotmaskinvare, ikke gjennom railwail API.

05

Spesifikasjoner

Modell-ID
octo-small
Utvikler
Other
Inndata
Tekst, Bilde
Utdata
Roboteraksjoner
Modellstørrelse
27M
Lisens
Apache-2.0 - fully open weights, code, and recipes. Research and commercial use permitted under the licence.
Katalogoppføring oppdatert
23. september 2026

Merkelapper

  • berkeley
  • vla
  • robotics
  • research-only
  • open-weights
  • small
  • consumer-gpu
06

Brukstilfeller

Hva det brukes til

  • Edge / on-robot inference on modest GPUs
  • Coursework and lab-class robot learning
  • Fast prototyping of new manipulation tasks
  • Small-budget academic research
  • Comparison baseline against larger VLAs
  • Starter checkpoint for new-robot fine-tuning
07

Ofte stilte spørsmål

Hva er Octo Small?

Octo Small er en modell fra Other i kategorien Robotikk / VLA. Den er oppført på Railwail, men kan ikke kjøres for øyeblikket.

Hvor mye koster Octo Small på Railwail?

Octo Small kan ikke kjøres på Railwail for øyeblikket, så det finnes ingen gjeldende pris. Tilgjengelige alternativer med priser er oppført lenger ned på denne siden.

Hvor rask er Octo Small?

Det finnes ennå ikke nok målte kjøringer av Octo Small på Railwail til å angi en kjøretid. Det avhenger av inndataene, innstillingene og belastningen hos leverandøren.

Når bør jeg bruke Octo Small?

Octo Small tilhører kategorien Robotikk / VLA. Kategorisiden viser de andre modellene av denne typen med prisene deres.

Alle modeller: Robotikk / VLA

Kan Octo Small behandle bilder?

Ja. Octo Small godtar bilder som inndata i tillegg til tekst.

Kan jeg bruke Octo Small akkurat nå?

Ikke tilgjengelig for øyeblikket. Siden forblir online; tilgjengelige alternativer fra samme kategori er oppført lenger ned.

Alle modeller gjennom én API

Én API-nøkkel for alle modeller på Railwail. Bruk belastes fra forhåndsbetalt kreditt, 1 kreditt = 0,01 USD.