Chatterbox
chatterboxResemble AI's open Chatterbox TTS. Zero-shot voice cloning from a short audio prompt with an exaggeration control for emotion intensity, plus CFG weight to balance pacing and fidelity.
- Price
- $0.030/1k chars
- Input → output
- Text → Audio
- Developer
- Resemble AI
- Updated
- September 23, 2026
Clone a voice
- 1
Your voice
Voice sample *Record up to 30 s · file up to 60 s, 10 MB10 to 30 seconds of clear speech, one speaker, no music or background noise.
Only your own voice or a voice you have verifiable permission for. Imitating other people without their consent is not allowed, see our Terms of Service.
- 2
Text
Example text, change it as you like.103 / 4,000 - 3
Listen
$0.0031 per run
Playground
Try Chatterbox
Input & output
This run
$0.0001 · 0.01 credits
New here?
10 free credits ($0.10) when you sign up with Google
Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits. Enough for 1,000 runs of this model.
Examples
- Length: 0:39
Prompt
We're excited to introduce Chatterbox, our first production-grade open source TTS model. Licensed under MIT, Chatterbox has been benchmarked against leading closed-source systems like ElevenLabs, and is consistently preferred in side-by-side evaluations. Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. It's also the first open source TTS model to support emotion exaggeration control, a powerful feature that makes your voices stand out. Try it now on our Hugging Face Gradio app. If you like the model but need to scale or finetune it for higher accuracy, check out our competitively priced TTS service (link). It delivers reliable performance with ultra-low latency of sub 200ms—ideal for production use in agents, applications, or interactive media.
- Length: 0:19
Prompt
Now let's make my mum's favourite. So three mars bars into the pan. Then we add the tuna and just stir for a bit, just let the chocolate and fish infuse. A sprinkle of olive oil and some tomato ketchup. Now smell that. Oh boy this is going to be incredible.
About Chatterbox
Chatterbox is a model by Resemble AI in the Text-to-speech category. On Railwail, Chatterbox costs $0.030 per 1,000 characters.
Pricing
| 1,000 characters | $0.030 per 1,000 characters |
|---|
- 1 credit = $0.01
Cost calculator
Price calculator
Total
$3.00
300 credits
Per run
$0.03 · 3 credits
Fixed price per run, known before the run starts.
API
No verified API example
The public API passes a different input format than this model needs. Use the playground above.
Specifications
- Model ID
chatterbox- Developer
- Resemble AI
- Category
- Text-to-speech
- Input
- Text
- Output
- Audio
- Billing
- Fixed price, known before the run
- Catalog entry updated
- September 23, 2026
Input parameters
Inputs and settings from the model's input schema. The example in the API section shows which of them the API accepts.
promptrequiredText to speak
Type: TextDefault: –Allowed values: up to 4,000 charactersseed0 = random
Type: IntegerDefault:0Allowed values: –audioOptional reference clip (5-60 s) to clone the voice from; requires consent
Type: –Default: –Allowed values: –cfg_weightType: NumberDefault:0.5Allowed values: 0.2 to 1temperatureType: NumberDefault:0.8Allowed values: 0.05 to 5exaggerationType: NumberDefault:0.5Allowed values: 0.25 to 2
Tags
- replicate
- resemble-ai
- tts
- voice-cloning
- expressive
Use cases
Frequently asked questions
What is Chatterbox?
Chatterbox is a model by Resemble AI in the Text-to-speech category.
How much does Chatterbox cost on Railwail?
On Railwail, Chatterbox costs $0.030 per 1,000 characters. The price is known before the run starts. Usage is paid from prepaid credits; 1 credit equals $0.01.
Which settings does Chatterbox support?
According to its input schema, Chatterbox knows these parameters: prompt (up to 4,000 characters), seed, audio, cfg_weight (0.2 to 1), temperature (0.05 to 5), and exaggeration (0.25 to 2).
How fast is Chatterbox?
There are not enough measured runs of Chatterbox on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.
Is Chatterbox better than AudioLDM 2?
That depends on the task. Chatterbox (Resemble AI) and AudioLDM 2 (Haohe Liu) are both models in the Text-to-speech category. The comparison page shows their prices and specifications side by side.
Compare Chatterbox and AudioLDM 2Comparable models
All in this category- OpenAI TTS-1OpenAI
OpenAI's text-to-speech model. Six built-in voices with natural intonation.
- OpenAI TTS-1 HDOpenAI
OpenAI's high-definition TTS model. Better quality for production use cases.
- Qwen3 TTSAlibaba (Qwen)
A unified Text-to-Speech demo featuring three powerful modes: Voice, Clone and Design
All models through one API
One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.