Incredibly Fast Whisper
incredibly-fast-whisperWhisper Large v3 wrapped with Hugging Face Transformers optimizations (batched inference, flash attention) for very high throughput. Transcribes hours of audio in minutes on a single GPU. Maintained by Vaibhav Srivastav. Good when you need bulk transcription fast.
- Price
- ≈ US$0.0056/run
- Input → output
- Audio → Text
- Developer
- Community
- Updated
- 23 September 2026
Playground
Try Incredibly Fast Whisper
Input & output
This run
about US$0.0056 · 0.56 credits
US$0.0166 (1.66 credits) are reserved at the start; the actual GPU time is billed.
New here?
10 free credits (US$0.10) when you sign up with Google
Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits. Enough for 6 runs of this model.
Examples
Prompt
transcribe
Response
the little tales they tell are false the door was barred locked and bolted as well ripe pears are fit hours fly by much too soon. The room was crowded with a mild wab. The room was crowded with a wild mob. This strong arm shall shield your honour. She blushed when he gave her a white orchid The beetle droned in the hot June sun
Prompt
transcribe
Response
院子门口不远处就是一个地铁站这是一个美丽而神奇的景象树上长满了又大又甜的桃子海豚和鲸鱼的表演是很好看的节目邮局门前的人行道上有一个蓝色的邮箱
Prompt
transcribe
Response
We have been a misunderstood and badly mocked org for a long time. When we started, we announced the org at the end of 2015 and said we were going to work on AGI. People thought we were batshit insane. I remember at the time, an eminent AI scientist at a large industrial AI lab was like DMing individual reporters being like, you know, these people aren't very good and it's ridiculous to talk about AGI and I can't believe you're giving them time of day. And it's like, that was the level of like pettiness and rancor in the field at a new group of people saying we're going to try to build AGI. So OpenAI and DeepMind was a small collection of folks who were brave enough to talk about AGI in the face of mockery. We don't get mocked as much now. Don't get mocked as much now. of OpenAI, the company behind GPT-4, JAD-GPT, DALI, Codex, and many other AI technologies, which both individually and together constitute some of the greatest breakthroughs in the history of artificial intelligence, computing, and humanity in general. Please allow me to say a few words about the possibilities and the dangers of AI in this current moment in the history of human civilization. I believe it is a critical moment. We stand on the precipice of fundamental societal transformation, where soon, nobody knows when, but many, including me, believe it's within our lifetime. The collective intelligence of the human species begins to pale in comparison, intelligence of the human species begins to pale in com…
About Incredibly Fast Whisper
Incredibly Fast Whisper is a model by Community in the Speech-to-text category. On Railwail, Incredibly Fast Whisper costs ≈ US$0.0056 per run.
Pricing
| Typical run (≈ 5 s on L40S) | US$0.0056 per run |
|---|---|
| GPU time (L40S) | US$0.00117 per GPU second |
- Billed by the GPU time the run actually takes. When the run starts, 3× the typical price is reserved from your balance and settled afterwards.
- 1 credit = US$0.01
Cost calculator
Price calculator
Typical according to the provider: about 4.7 s
Total
US$0.56
56 credits
Per run
US$0.0056 · 0.56 credits
Billed by the actual GPU time; this is an estimate.
API
curl https://railwail.com/api/v1/audio/transcriptions \
-H "Authorization: Bearer $RAILWAIL_API_KEY" \
-F model='incredibly-fast-whisper' \
-F file=@audio.mp3import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["RAILWAIL_API_KEY"],
base_url="https://railwail.com/api/v1",
)
with open("audio.mp3", "rb") as audio:
transcript = client.audio.transcriptions.create(
model="incredibly-fast-whisper",
file=audio,
)
print(transcript.text)import fs from "node:fs";
import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.RAILWAIL_API_KEY,
baseURL: "https://railwail.com/api/v1",
});
const transcript = await client.audio.transcriptions.create({
model: "incredibly-fast-whisper",
file: fs.createReadStream("audio.mp3"),
});
console.log(transcript.text);Specifications
- Model ID
incredibly-fast-whisper- Developer
- Community
- Category
- Speech-to-text
- Input
- Audio
- Output
- Text
- Billing
- By usage (tokens or GPU time)
- Catalog entry updated
- 23 September 2026
Input parameters
Inputs and settings from the model's input schema. The example in the API section shows which of them the API accepts.
audiorequiredURL or upload of the audio file to transcribe
Type: TextDefault: –Allowed values: –taskType: ChoiceDefault:transcribeAllowed values: transcribe or translatelanguageOptional language hint (e.g. english, german); auto-detected if omitted
Type: TextDefault: –Allowed values: –timestampType: ChoiceDefault:chunkAllowed values: chunk or wordbatch_sizeType: IntegerDefault:24Allowed values: 1 to 64
Tags
- replicate
- whisper
- stt
- transcription
- fast
- batched
- multilingual
Use cases
Frequently asked questions
What is Incredibly Fast Whisper?
Incredibly Fast Whisper is a model by Community in the Speech-to-text category. On Railwail you can call it with an API key through the Railwail API.
How much does Incredibly Fast Whisper cost on Railwail?
On Railwail, Incredibly Fast Whisper costs ≈ US$0.0056 per run. You are charged for what each request actually uses. Usage is paid from prepaid credits; 1 credit equals US$0.01.
Which settings does Incredibly Fast Whisper support?
According to its input schema, Incredibly Fast Whisper knows these parameters: audio, task (transcribe or translate), language, timestamp (chunk or word) and batch_size (1 to 64).
How fast is Incredibly Fast Whisper?
There are not enough measured runs of Incredibly Fast Whisper on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.
Is Incredibly Fast Whisper better than Whisper?
That depends on the task. Incredibly Fast Whisper (Community) and Whisper (OpenAI) are both models in the Speech-to-text category. The comparison page shows their prices and specifications side by side.
Compare Incredibly Fast Whisper and WhisperHow do I use Incredibly Fast Whisper through the API?
Create a Railwail API key and send your request with the model ID incredibly-fast-whisper. Code examples for curl, Python and JavaScript are in the API section of this page.
Comparable models
All in this category- WhisperOpenAI
OpenAI's Whisper running on Replicate. General-purpose speech recognition trained on 680k hours of multilingual audio. Transcribes and translates 99 languages, robust to accents and background noise, and outputs plain text, segments, or word-level timestamps.
- SeamlessM4TCommunity
Meta's SeamlessM4T multimodal translation model. Takes speech or text input and produces transcription or translation across about 100 languages, including speech-to-text and speech-to-speech. One model covers ASR plus cross-lingual translation without chaining separate systems.
- SeamlessM4T v2 Large (Speech)Community
Meta SeamlessM4T v2 Large speech mode. Speech-to-speech, speech-to-text, and text-to-speech translation across 100+ languages in a single unified model.
≈ US$0.0012/run
79 % cheaper per unit
Compare Incredibly Fast Whisper vs. SeamlessM4T v2 Large (Speech)
Use Incredibly Fast Whisper via the API
One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = US$0.01.