Whisper Large v3 Turbo

Riconoscimento vocale (STT)Non disponibile
di OpenAIID modello: whisper-large-v3-turbo

OpenAI's distilled Whisper Large v3. ~216x realtime, 99+ languages, MIT-licensed weights.

Stato
Non disponibile
Input → output
Audio → Testo
Sviluppatore
OpenAI
Aggiornato
23 settembre 2026

Whisper Large v3 Turbo non è attualmente disponibile

Attualmente non disponibile: questo modello è stato disattivato.

Puoi comunque leggere i dettagli su questa pagina. Scegli una delle alternative disponibili di seguito per eseguire subito un modello comparabile.

Vai alle alternative
01

Modelli comparabili

Tutti in questa categoria
  • Whisper Large v3 wrapped with Hugging Face Transformers optimizations (batched inference, flash attention) for very high throughput. Transcribes hours of audio in minutes on a single GPU. Maintained by Vaibhav Srivastav. Good when you need bulk transcription fast.

    ≈ 0,0056 USD/esecuzione

  • WhisperOpenAI

    OpenAI's Whisper running on Replicate. General-purpose speech recognition trained on 680k hours of multilingual audio. Transcribes and translates 99 languages, robust to accents and background noise, and outputs plain text, segments, or word-level timestamps.

    ≈ 0,0034 USD/esecuzione

  • SeamlessM4TCommunity

    Meta's SeamlessM4T multimodal translation model. Takes speech or text input and produces transcription or translation across about 100 languages, including speech-to-text and speech-to-speech. One model covers ASR plus cross-lingual translation without chaining separate systems.

    ≈ 0,156 USD/esecuzione

02

Playground

Prova Whisper Large v3 Turbo

Nessuna maschera di input

Attualmente non disponibile

Attualmente non disponibile: questo modello è stato disattivato.

Il playground è disabilitato. Trovi modelli comparabili nella stessa categoria: Visualizza alternative

03

Informazioni su Whisper Large v3 Turbo

RiassuntoA partire da 23 settembre 2026

Whisper Large v3 Turbo è un modello di OpenAI nella categoria Riconoscimento vocale (STT). Whisper Large v3 Turbo non è attualmente disponibile su Railwail.

Sfondo

Informazioni su OpenAI

Fondato 2015 · San Francisco, California, USA

OpenAI was founded in December 2015 by Sam Altman, Elon Musk, Greg Brockman, Ilya Sutskever, Wojciech Zaremba and John Schulman, restructured to capped-profit OpenAI LP in 2019. Whisper Large v3 Turbo was released in October 2024 as a distilled fast variant of Whisper Large v3, designed to deliver approximately 8x faster inference at near-identical accuracy by reducing the decoder depth from 32 to 4 layers. The release was led by the original Whisper authors (Alec Radford, Jong Wook Kim, Tao Xu) and remained under the MIT licence. Turbo was distributed via GitHub, the Hugging Face Hub and the OpenAI Whisper API as the new default model where supported, replacing many in-production deployments of Whisper Large v3 within weeks of launch.

Visita OpenAI

Architettura

Distilled encoder-decoder Transformer (4-layer decoder) for speech recognition

Whisper Large v3 Turbo is a distilled variant of Whisper Large v3 that keeps the same 32-layer audio encoder and 128-mel front-end but shrinks the decoder from 32 to just 4 Transformer layers, taking the total parameter count from 1.55B to 809M. The smaller decoder gives roughly 8x faster inference on long-form audio (and 4-5x faster on short clips) at a WER cost of approximately 0.5-1 percentage points on most benchmarks. The model was distilled on the same multilingual corpus as Large v3 (5 million hours total, of which 4 million are pseudo-labelled) with knowledge-distillation losses from the Large v3 teacher. Translation-to-English capability was deliberately removed to focus capacity on transcription quality. The 30-second sliding window, 99-language coverage and special task tokens are unchanged. Turbo runs in real-time on consumer GPUs (RTX 3060) and at 3-4x real-time on Apple Silicon CPUs via whisper.cpp.

Parametri
809M
Contesto
30 token

Capacità

  • 8x faster long-form transcription than Whisper Large v3
  • 99-language transcription with automatic language detection
  • Word-level timestamps preserved
  • Runs in real-time on a single consumer GPU (RTX 3060 / M2 Pro)
  • Half the memory footprint of Large v3 (809M vs 1.55B)
  • Open weights under MIT licence
  • Drop-in replacement for Large v3 in most pipelines
  • Best for: production ASR on commodity hardware, on-premise transcription, batch processing

Addestramento e licenza

Distilled from Whisper Large v3 on the same 5-million-hour multilingual audio corpus with knowledge-distillation losses. Translation-to-English data was excluded.

Licenza: MIT licence for code and weights; commercial use permitted.

Test di sicurezza: Inherits all hallucination and bias caveats from Whisper Large v3; no separate red-team report.

Limitazioni note

  • No translation-to-English mode (transcription only)
  • WER 0.5-1 pp worse than Large v3 on average
  • Same 30-second hard window requires chunking
  • Same hallucination behaviour on silent / music-only audio
  • No native diarisation
04

Prezzi

Attualmente non disponibile: questo modello è stato disattivato. Al momento non c'è un prezzo per questo modello, quindi non può essere eseguito.

05

API

Chiama Whisper Large v3 Turbo con la tua chiave API Railwail. Usa questo ID modello nella richiesta:

Attualmente non disponibile

Il modello non ha un prezzo verificato o è disattivato; le chiamate API vengono rifiutate.

06

Specifiche

ID modello
whisper-large-v3-turbo
Sviluppatore
OpenAI
Input
Audio
Output
Testo
Formati di output
JSON, SRT, VTT
Ciclo di vita
Non disponibile
Dimensione del modello
809M
Licenza
MIT licence for code and weights; commercial use permitted.
Voce di catalogo aggiornata
23 settembre 2026

Parametri di input

Input e impostazioni dallo schema di input del modello. L'esempio nella sezione API mostra quali di essi l'API accetta.

  • fileObbligatorio

    URL or upload path to audio file

    Tipo: Testo
    Predefinito: –
    Valori consentiti: –
  • prompt

    Optional context to guide transcription

    Tipo: Testo
    Predefinito: –
    Valori consentiti: fino a 1000 caratteri
  • language

    Optional ISO-639-1 language code (e.g. en, de, fr)

    Tipo: Testo
    Predefinito: –
    Valori consentiti: –
  • temperature
    Tipo: Numero
    Predefinito: 0
    Valori consentiti: 0 a 1
  • response_format
    Tipo: Scelta
    Predefinito: json
    Valori consentiti: json, text, srt o vtt

Etichette

  • openai
  • whisper
  • stt
  • transcription
  • open-weights
  • multilingual
  • per-minute
07

Casi d'uso

A cosa serve

  • Real-time transcription on consumer GPUs
  • Batch transcription of large podcast / lecture archives
  • On-device transcription via whisper.cpp
  • Cost-sensitive production ASR pipelines
  • Edge deployments where bandwidth is limited
08

Domande frequenti

Cos'è Whisper Large v3 Turbo?

Whisper Large v3 Turbo è un modello di OpenAI nella categoria Riconoscimento vocale (STT). È elencato su Railwail ma non può essere eseguito al momento.

Quanto costa Whisper Large v3 Turbo su Railwail?

Whisper Large v3 Turbo non può essere eseguito su Railwail al momento, quindi non c'è un prezzo attuale. Le alternative disponibili con i prezzi sono elencate più in basso in questa pagina.

Quali impostazioni supporta Whisper Large v3 Turbo?

Secondo il suo schema di input, Whisper Large v3 Turbo conosce questi parametri: file, prompt (fino a 1000 caratteri), language, temperature (0 a 1) e response_format (json, text, srt o vtt).

Quanto è veloce Whisper Large v3 Turbo?

Non ci sono ancora abbastanza esecuzioni misurate di Whisper Large v3 Turbo su Railwail per indicare un tempo di esecuzione. Dipende dall'input, dalle impostazioni e dal carico presso il provider.

Whisper Large v3 Turbo è migliore di Incredibly Fast Whisper?

Dipende dall'attività. Whisper Large v3 Turbo (OpenAI) e Incredibly Fast Whisper (Community) sono entrambi modelli nella categoria Riconoscimento vocale (STT). La pagina di confronto mostra i loro prezzi e le specifiche affiancati.

Confronta Whisper Large v3 Turbo e Incredibly Fast Whisper

Posso usare Whisper Large v3 Turbo adesso?

Attualmente non disponibile: questo modello è stato disattivato. La pagina rimane online; le alternative disponibili della stessa categoria sono elencate più in basso.

Tutti i modelli tramite un'API

Una chiave API per tutti i modelli su Railwail. L'utilizzo viene addebitato da crediti prepagati, 1 credito = 0,01 USD.