AI speech, transcription and music models by price

Text to speech, transcription and music models, by price. The Arena dataset has no arena for these models, so there is no quality or value ranking yet.

AI speech, transcription and music models by price

Sorted by the price of one default run, lowest first.

Text to speech · price per 1,000 characters

AI speech, transcription and music models by price · Text to speech · price per 1,000 characters
01$0.018/1k charsTry
02
Qwen3 TTS
Alibaba (Qwen)
$0.024/1k charsTry
03
Chatterbox
Resemble AI
$0.030/1k charsTry
04$0.036/1k charsTry

Text to speech · price per default run

AI speech, transcription and music models by price · Text to speech · price per default run
01
AudioLDM 2
Haohe Liu
≈ $0.0157/runTry
02
F5-TTS
X-LANCE (SJTU)
≈ $0.0168/runTry
03
Riffusion
Riffusion
≈ $0.0576/runTry
04≈ $0.0972/runTry

Speech to text

AI speech, transcription and music models by price · Speech to text
01
Whisper
OpenAI
≈ $0.0034/runTry

Music & audio

AI speech, transcription and music models by price · Music & audio
01≈ $0.0529/runTry

How to read this ranking

Price
What one run with the model’s default settings costs on Railwail, from the same pricing rules that bill your runs (1 credit = $0.01). ≈ marks GPU-time prices: a typical run, billed by the real run time.
Who is listed
Every model of this category you can run on Railwail right now.
Rank
The position within this list, not the rank on the arena.

Questions about this ranking

Why is there no quality ranking for audio models?

The Arena leaderboard dataset has no arena for speech, transcription or music models. The list shows what we can state exactly: the price of each model on Railwail.

AI speech, transcription and music models by price | Railwail