Whisper Large v3 Turbo

Spraak-naar-tekstNiet beschikbaar
van OpenAIModel-ID: whisper-large-v3-turbo

OpenAI's distilled Whisper Large v3. ~216x realtime, 99+ languages, MIT-licensed weights.

Status
Niet beschikbaar
Invoer → Uitvoer
Audio → Tekst
Ontwikkelaar
OpenAI
Bijgewerkt
23 september 2026

Whisper Large v3 Turbo is momenteel niet beschikbaar

Momenteel niet beschikbaar: dit model is gedeactiveerd.

Je kunt de details op deze pagina nog steeds lezen. Kies een van de beschikbare alternatieven hieronder om direct een vergelijkbaar model uit te voeren.

Naar alternatieven
01

Vergelijkbare modellen

Alle in deze categorie
  • Whisper Large v3 wrapped with Hugging Face Transformers optimizations (batched inference, flash attention) for very high throughput. Transcribes hours of audio in minutes on a single GPU. Maintained by Vaibhav Srivastav. Good when you need bulk transcription fast.

    ≈ US$ 0,0056/uitvoering

  • WhisperOpenAI

    OpenAI's Whisper running on Replicate. General-purpose speech recognition trained on 680k hours of multilingual audio. Transcribes and translates 99 languages, robust to accents and background noise, and outputs plain text, segments, or word-level timestamps.

    ≈ US$ 0,0034/uitvoering

  • SeamlessM4TCommunity

    Meta's SeamlessM4T multimodal translation model. Takes speech or text input and produces transcription or translation across about 100 languages, including speech-to-text and speech-to-speech. One model covers ASR plus cross-lingual translation without chaining separate systems.

    ≈ US$ 0,156/uitvoering

02

Playground

Whisper Large v3 Turbo proberen

Geen invoerformulier

Momenteel niet beschikbaar

Momenteel niet beschikbaar: dit model is gedeactiveerd.

De playground is uitgeschakeld. Vergelijkbare modellen vind je in dezelfde categorie: Alternatieven bekijken

03

Over Whisper Large v3 Turbo

SamengevatPer 23 september 2026

Whisper Large v3 Turbo is een model van OpenAI in de categorie Spraak-naar-tekst. Whisper Large v3 Turbo is momenteel niet beschikbaar op Railwail.

Achtergrond

Over OpenAI

Opgericht 2015 · San Francisco, California, USA

OpenAI was founded in December 2015 by Sam Altman, Elon Musk, Greg Brockman, Ilya Sutskever, Wojciech Zaremba and John Schulman, restructured to capped-profit OpenAI LP in 2019. Whisper Large v3 Turbo was released in October 2024 as a distilled fast variant of Whisper Large v3, designed to deliver approximately 8x faster inference at near-identical accuracy by reducing the decoder depth from 32 to 4 layers. The release was led by the original Whisper authors (Alec Radford, Jong Wook Kim, Tao Xu) and remained under the MIT licence. Turbo was distributed via GitHub, the Hugging Face Hub and the OpenAI Whisper API as the new default model where supported, replacing many in-production deployments of Whisper Large v3 within weeks of launch.

OpenAI bezoeken

Architectuur

Distilled encoder-decoder Transformer (4-layer decoder) for speech recognition

Whisper Large v3 Turbo is a distilled variant of Whisper Large v3 that keeps the same 32-layer audio encoder and 128-mel front-end but shrinks the decoder from 32 to just 4 Transformer layers, taking the total parameter count from 1.55B to 809M. The smaller decoder gives roughly 8x faster inference on long-form audio (and 4-5x faster on short clips) at a WER cost of approximately 0.5-1 percentage points on most benchmarks. The model was distilled on the same multilingual corpus as Large v3 (5 million hours total, of which 4 million are pseudo-labelled) with knowledge-distillation losses from the Large v3 teacher. Translation-to-English capability was deliberately removed to focus capacity on transcription quality. The 30-second sliding window, 99-language coverage and special task tokens are unchanged. Turbo runs in real-time on consumer GPUs (RTX 3060) and at 3-4x real-time on Apple Silicon CPUs via whisper.cpp.

Parameters
809M
Context
30 tokens

Mogelijkheden

  • 8x faster long-form transcription than Whisper Large v3
  • 99-language transcription with automatic language detection
  • Word-level timestamps preserved
  • Runs in real-time on a single consumer GPU (RTX 3060 / M2 Pro)
  • Half the memory footprint of Large v3 (809M vs 1.55B)
  • Open weights under MIT licence
  • Drop-in replacement for Large v3 in most pipelines
  • Best for: production ASR on commodity hardware, on-premise transcription, batch processing

Training & licentie

Distilled from Whisper Large v3 on the same 5-million-hour multilingual audio corpus with knowledge-distillation losses. Translation-to-English data was excluded.

Licentie: MIT licence for code and weights; commercial use permitted.

Veiligheidstests: Inherits all hallucination and bias caveats from Whisper Large v3; no separate red-team report.

Bekende beperkingen

  • No translation-to-English mode (transcription only)
  • WER 0.5-1 pp worse than Large v3 on average
  • Same 30-second hard window requires chunking
  • Same hallucination behaviour on silent / music-only audio
  • No native diarisation
04

Prijzen

Momenteel niet beschikbaar: dit model is gedeactiveerd. Er is momenteel geen prijs voor dit model, dus het kan niet worden uitgevoerd.

05

API

Roep Whisper Large v3 Turbo aan met je Railwail API-sleutel. Gebruik deze model-ID in het verzoek:

Momenteel niet beschikbaar

Het model heeft geen geverifieerde prijs of is gedeactiveerd; API-aanroepen worden geweigerd.

06

Specificaties

Model-ID
whisper-large-v3-turbo
Ontwikkelaar
OpenAI
Invoer
Audio
Uitvoer
Tekst
Uitvoerformaten
JSON, SRT, VTT
Levenscyclus
Niet beschikbaar
Modelgrootte
809M
Licentie
MIT licence for code and weights; commercial use permitted.
Catalogusitem bijgewerkt
23 september 2026

Invoerparameters

Invoeren en instellingen uit het invoerschema van het model. Het voorbeeld in de API-sectie toont welke daarvan de API accepteert.

  • fileVerplicht

    URL or upload path to audio file

    Type: Tekst
    Standaard: –
    Toegestane waarden: –
  • prompt

    Optional context to guide transcription

    Type: Tekst
    Standaard: –
    Toegestane waarden: tot 1.000 tekens
  • language

    Optional ISO-639-1 language code (e.g. en, de, fr)

    Type: Tekst
    Standaard: –
    Toegestane waarden: –
  • temperature
    Type: Getal
    Standaard: 0
    Toegestane waarden: 0 tot 1
  • response_format
    Type: Keuze
    Standaard: json
    Toegestane waarden: json, text, srt of vtt

Tags

  • openai
  • whisper
  • stt
  • transcription
  • open-weights
  • multilingual
  • per-minute
07

Gebruiksscenario's

Waarvoor het wordt gebruikt

  • Real-time transcription on consumer GPUs
  • Batch transcription of large podcast / lecture archives
  • On-device transcription via whisper.cpp
  • Cost-sensitive production ASR pipelines
  • Edge deployments where bandwidth is limited
08

Veelgestelde vragen

Wat is Whisper Large v3 Turbo?

Whisper Large v3 Turbo is een model van OpenAI in de categorie Spraak-naar-tekst. Het staat in de Railwail-catalogus, maar kan momenteel niet worden uitgevoerd.

Hoeveel kost Whisper Large v3 Turbo op Railwail?

Whisper Large v3 Turbo kan momenteel niet op Railwail worden uitgevoerd, dus er is geen huidige prijs. Beschikbare alternatieven met prijzen staan verderop op deze pagina.

Welke instellingen ondersteunt Whisper Large v3 Turbo?

Volgens het invoerschema kent Whisper Large v3 Turbo deze parameters: file, prompt (tot 1.000 tekens), language, temperature (0 tot 1) en response_format (json, text, srt of vtt).

Hoe snel is Whisper Large v3 Turbo?

Er zijn nog niet genoeg gemeten runs van Whisper Large v3 Turbo op Railwail om een uitvoeringstijd op te geven. Dit hangt af van de invoer, de instellingen en de belasting bij de provider.

Is Whisper Large v3 Turbo beter dan Incredibly Fast Whisper?

Dat hangt van de taak af. Whisper Large v3 Turbo (OpenAI) en Incredibly Fast Whisper (Community) zijn beide modellen in de categorie Spraak-naar-tekst. De vergelijkingspagina toont hun prijzen en specificaties naast elkaar.

Whisper Large v3 Turbo en Incredibly Fast Whisper vergelijken

Kan ik Whisper Large v3 Turbo nu gebruiken?

Momenteel niet beschikbaar: dit model is gedeactiveerd. De pagina blijft online; beschikbare alternatieven uit dezelfde categorie staan verderop.

Alle modellen via één API

Één API-sleutel voor elk model op Railwail. Gebruik wordt afgerekend via vooraf gekochte credits, 1 credit = US$ 0,01.