Whisper Large v3 Turbo

Konverzia reči na textNedostupné
od OpenAIID modelu: whisper-large-v3-turbo

OpenAI's distilled Whisper Large v3. ~216x realtime, 99+ languages, MIT-licensed weights.

Stav
Nedostupné
Vstup → výstup
Audio → Text
Vývojár
OpenAI
Aktualizované
23. septembra 2026

Whisper Large v3 Turbo nie je momentálne dostupný

Momentálne nedostupné: tento model bol deaktivovaný.

Podrobnosti na tejto stránke si môžete prečítať. Vyberte si jednu z dostupných alternatív nižšie a spustite porovnateľný model hneď.

Prejsť na alternatívy
01

Porovnateľné modely

Všetky v tejto kategórii
  • Whisper Large v3 wrapped with Hugging Face Transformers optimizations (batched inference, flash attention) for very high throughput. Transcribes hours of audio in minutes on a single GPU. Maintained by Vaibhav Srivastav. Good when you need bulk transcription fast.

    ≈ 0,0056 USD/spustenie

  • WhisperOpenAI

    OpenAI's Whisper running on Replicate. General-purpose speech recognition trained on 680k hours of multilingual audio. Transcribes and translates 99 languages, robust to accents and background noise, and outputs plain text, segments, or word-level timestamps.

    ≈ 0,0034 USD/spustenie

  • SeamlessM4TCommunity

    Meta's SeamlessM4T multimodal translation model. Takes speech or text input and produces transcription or translation across about 100 languages, including speech-to-text and speech-to-speech. One model covers ASR plus cross-lingual translation without chaining separate systems.

    ≈ 0,156 USD/spustenie

02

Playground

Vyskúšajte Whisper Large v3 Turbo

Bez vstupného formulára

Momentálne nedostupné

Momentálne nedostupné: tento model bol deaktivovaný.

Playground je vypnutý. Porovnateľné modely nájdete v tej istej kategórii: Pozrieť alternatívy

03

O Whisper Large v3 Turbo

StručneK 23. septembra 2026

Whisper Large v3 Turbo je model od OpenAI v kategórii Konverzia reči na text. Whisper Large v3 Turbo nie je v súčasnosti dostupný na Railwail.

Pozadie

O OpenAI

Založené 2015 · San Francisco, California, USA

OpenAI was founded in December 2015 by Sam Altman, Elon Musk, Greg Brockman, Ilya Sutskever, Wojciech Zaremba and John Schulman, restructured to capped-profit OpenAI LP in 2019. Whisper Large v3 Turbo was released in October 2024 as a distilled fast variant of Whisper Large v3, designed to deliver approximately 8x faster inference at near-identical accuracy by reducing the decoder depth from 32 to 4 layers. The release was led by the original Whisper authors (Alec Radford, Jong Wook Kim, Tao Xu) and remained under the MIT licence. Turbo was distributed via GitHub, the Hugging Face Hub and the OpenAI Whisper API as the new default model where supported, replacing many in-production deployments of Whisper Large v3 within weeks of launch.

Navštíviť OpenAI

Architektúra

Distilled encoder-decoder Transformer (4-layer decoder) for speech recognition

Whisper Large v3 Turbo is a distilled variant of Whisper Large v3 that keeps the same 32-layer audio encoder and 128-mel front-end but shrinks the decoder from 32 to just 4 Transformer layers, taking the total parameter count from 1.55B to 809M. The smaller decoder gives roughly 8x faster inference on long-form audio (and 4-5x faster on short clips) at a WER cost of approximately 0.5-1 percentage points on most benchmarks. The model was distilled on the same multilingual corpus as Large v3 (5 million hours total, of which 4 million are pseudo-labelled) with knowledge-distillation losses from the Large v3 teacher. Translation-to-English capability was deliberately removed to focus capacity on transcription quality. The 30-second sliding window, 99-language coverage and special task tokens are unchanged. Turbo runs in real-time on consumer GPUs (RTX 3060) and at 3-4x real-time on Apple Silicon CPUs via whisper.cpp.

Parametre
809M
Kontext
30 tokenov

Schopnosti

  • 8x faster long-form transcription than Whisper Large v3
  • 99-language transcription with automatic language detection
  • Word-level timestamps preserved
  • Runs in real-time on a single consumer GPU (RTX 3060 / M2 Pro)
  • Half the memory footprint of Large v3 (809M vs 1.55B)
  • Open weights under MIT licence
  • Drop-in replacement for Large v3 in most pipelines
  • Best for: production ASR on commodity hardware, on-premise transcription, batch processing

Tréning a licencia

Distilled from Whisper Large v3 on the same 5-million-hour multilingual audio corpus with knowledge-distillation losses. Translation-to-English data was excluded.

Licencia: MIT licence for code and weights; commercial use permitted.

Bezpečnostné testy: Inherits all hallucination and bias caveats from Whisper Large v3; no separate red-team report.

Známe obmedzenia

  • No translation-to-English mode (transcription only)
  • WER 0.5-1 pp worse than Large v3 on average
  • Same 30-second hard window requires chunking
  • Same hallucination behaviour on silent / music-only audio
  • No native diarisation
04

Ceny

Momentálne nedostupné: tento model bol deaktivovaný. V súčasnosti nie je cena za tento model, preto ho nie je možné spustiť.

05

API

Zavolajte Whisper Large v3 Turbo s vaším API kľúčom Railwail. V požiadavke použite toto ID modelu:

Momentálne nedostupné

Model nemá overenú cenu alebo je deaktivovaný; volania API sú odmietnuté.

06

Špecifikácie

ID modelu
whisper-large-v3-turbo
Vývojár
OpenAI
Vstup
Audio
Výstup
Text
Formáty výstupu
JSON, SRT, VTT
Životný cyklus
Nedostupné
Veľkosť modelu
809M
Licencia
MIT licence for code and weights; commercial use permitted.
Katalógová položka aktualizovaná
23. septembra 2026

Vstupné parametre

Vstupy a nastavenia zo vstupnej schémy modelu. Príklad v sekcii API ukazuje, ktoré z nich API akceptuje.

  • filepovinné

    URL or upload path to audio file

    Typ: Text
    Predvolené: –
    Povolené hodnoty: –
  • prompt

    Optional context to guide transcription

    Typ: Text
    Predvolené: –
    Povolené hodnoty: až 1 000 znakov
  • language

    Optional ISO-639-1 language code (e.g. en, de, fr)

    Typ: Text
    Predvolené: –
    Povolené hodnoty: –
  • temperature
    Typ: Číslo
    Predvolené: 0
    Povolené hodnoty: 0 až 1
  • response_format
    Typ: Výber
    Predvolené: json
    Povolené hodnoty: json, text, srt alebo vtt

Značky

  • openai
  • whisper
  • stt
  • transcription
  • open-weights
  • multilingual
  • per-minute
07

Prípady použitia

Na čo sa používa

  • Real-time transcription on consumer GPUs
  • Batch transcription of large podcast / lecture archives
  • On-device transcription via whisper.cpp
  • Cost-sensitive production ASR pipelines
  • Edge deployments where bandwidth is limited
08

Často kladené otázky

Čo je Whisper Large v3 Turbo?

Whisper Large v3 Turbo je model od OpenAI v kategórii Konverzia reči na text. Je uvedený na Railwail, ale momentálne ho nie je možné spustiť.

Koľko stojí Whisper Large v3 Turbo na Railwail?

Whisper Large v3 Turbo nie je momentálne možné spustiť na Railwail, preto nie je aktuálna cena. Dostupné alternatívy s cenami sú uvedené nižšie na tejto stránke.

Ktoré nastavenia podporuje Whisper Large v3 Turbo?

Podľa svojej vstupnej schémy Whisper Large v3 Turbo pozná tieto parametre: file, prompt (až 1 000 znakov), language, temperature (0 až 1) a response_format (json, text, srt alebo vtt).

Ako rýchly je Whisper Large v3 Turbo?

Pre Whisper Large v3 Turbo je na Railwail zatiaľ príliš málo meraných spustení na určenie doby spustenia. Závisí to od vstupu, nastavení a zaťaženia u poskytovateľa.

Je Whisper Large v3 Turbo lepší ako Incredibly Fast Whisper?

Závisí to od úlohy. Whisper Large v3 Turbo (OpenAI) a Incredibly Fast Whisper (Community) sú oba modely v kategórii Konverzia reči na text. Stránka porovnania zobrazuje ich ceny a špecifikácie vedľa seba.

Porovnať Whisper Large v3 Turbo a Incredibly Fast Whisper

Môžem Whisper Large v3 Turbo používať práve teraz?

Momentálne nedostupné: tento model bol deaktivovaný. Stránka zostáva online; dostupné alternatívy z tej istej kategórie sú uvedené nižšie.

Všetky modely cez jedno API

Jeden API kľúč pre všetky modely na Railwail. Použitie sa účtuje z predplateného kreditu, 1 kredit = 0,01 USD.