AI speech, transcription and music models by price
Text to speech, transcription and music models, by price. The Arena dataset has no arena for these models, so there is no quality or value ranking yet.
AI speech, transcription and music models by price
Sorted by the price of one default run, lowest first.
Text to speech · price per 1,000 characters
| # | Model | Price | Actions |
|---|---|---|---|
| 01 | OpenAI TTS-1 OpenAI | US$0.018/1k chars | Try |
| 02 | Qwen3 TTS Alibaba (Qwen) | US$0.024/1k chars | Try |
| 03 | Chatterbox Resemble AI | US$0.030/1k chars | Try |
| 04 | OpenAI TTS-1 HD OpenAI | US$0.036/1k chars | Try |
Text to speech · price per default run
Speech to text
How to read this ranking
- Price
- What one run with the model’s default settings costs on Railwail, from the same pricing rules that bill your runs (1 credit = $0.01). ≈ marks GPU-time prices: a typical run, billed by the real run time.
- Who is listed
- Every model of this category you can run on Railwail right now.
- Rank
- The position within this list, not the rank on the arena.
Questions about this ranking
Why is there no quality ranking for audio models?
The Arena leaderboard dataset has no arena for speech, transcription or music models. The list shows what we can state exactly: the price of each model on Railwail.