Chatterbox
chatterboxResemble AI's open Chatterbox TTS. Zero-shot voice cloning from a short audio prompt with an exaggeration control for emotion intensity, plus CFG weight to balance pacing and fidelity.
- Prezzo
- 0,030Â USD/1.000 caratteri
- Input → output
- Testo → Audio
- Sviluppatore
- Resemble AI
- Aggiornato
- 23 settembre 2026
Clone a voice
- 1
Your voice
Voice sample *Record up to 30 s · file up to 60 s, 10 MB10 to 30 seconds of clear speech, one speaker, no music or background noise.
Only your own voice or a voice you have verifiable permission for. Imitating other people without their consent is not allowed, see our Terms of Service.
- 2
Text
Example text, change it as you like.103 / 4000 - 3
Listen
0,0031Â USD per run
Playground
Prova Chatterbox
Input e output
Questa esecuzione
0,0001 USD · 0,01 crediti
Nuovo qui?
10 crediti gratuiti (0,10Â USD) quando ti iscrivi con Google
Utilizzabile 24 ore dopo l'iscrizione, fino a 5 esecuzioni al giorno e al massimo 2 crediti per esecuzione. Altri metodi di accesso iniziano senza crediti. Sufficiente per 1000 esecuzioni di questo modello.
Examples
- Durata: 0:39
Prompt
We're excited to introduce Chatterbox, our first production-grade open source TTS model. Licensed under MIT, Chatterbox has been benchmarked against leading closed-source systems like ElevenLabs, and is consistently preferred in side-by-side evaluations. Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. It's also the first open source TTS model to support emotion exaggeration control, a powerful feature that makes your voices stand out. Try it now on our Hugging Face Gradio app. If you like the model but need to scale or finetune it for higher accuracy, check out our competitively priced TTS service (link). It delivers reliable performance with ultra-low latency of sub 200ms—ideal for production use in agents, applications, or interactive media.
- Durata: 0:19
Prompt
Now let's make my mum's favourite. So three mars bars into the pan. Then we add the tuna and just stir for a bit, just let the chocolate and fish infuse. A sprinkle of olive oil and some tomato ketchup. Now smell that. Oh boy this is going to be incredible.
Informazioni su Chatterbox
Chatterbox è un modello di Resemble AI nella categoria Sintesi vocale (TTS). Su Railwail, Chatterbox costa 0,030 USD per 1.000 caratteri.
Prezzi
| 1.000 caratteri | 0,030Â USD per 1.000 caratteri |
|---|
- 1 credito = 0,01Â USD
Calcolatore di costi
Calcolatore prezzi
Totale
3,00Â USD
300 crediti
Per esecuzione
0,03 USD · 3 crediti
Prezzo fisso per esecuzione, noto prima dell'inizio.
API
Nessun esempio API verificato
L'API pubblica passa un formato di input diverso da quello richiesto da questo modello. Usa il playground sopra.
Specifiche
- ID modello
chatterbox- Sviluppatore
- Resemble AI
- Categoria
- Sintesi vocale (TTS)
- Input
- Testo
- Output
- Audio
- Fatturazione
- Prezzo fisso, noto prima dell'esecuzione
- Voce di catalogo aggiornata
- 23 settembre 2026
Parametri di input
Input e impostazioni dallo schema di input del modello. L'esempio nella sezione API mostra quali di essi l'API accetta.
promptObbligatorioText to speak
Tipo: TestoPredefinito: –Valori consentiti: fino a 4000 caratteriseed0 = random
Tipo: Numero interoPredefinito:0Valori consentiti: –audioOptional reference clip (5-60 s) to clone the voice from; requires consent
Tipo: –Predefinito: –Valori consentiti: –cfg_weightTipo: NumeroPredefinito:0.5Valori consentiti: 0,2 a 1temperatureTipo: NumeroPredefinito:0.8Valori consentiti: 0,05 a 5exaggerationTipo: NumeroPredefinito:0.5Valori consentiti: 0,25 a 2
Etichette
- replicate
- resemble-ai
- tts
- voice-cloning
- expressive
Casi d'uso
Domande frequenti
Cos'è Chatterbox?
Chatterbox è un modello di Resemble AI nella categoria Sintesi vocale (TTS).
Quanto costa Chatterbox su Railwail?
Su Railwail, Chatterbox costa 0,030 USD per 1.000 caratteri. Il prezzo è noto prima dell'inizio dell'esecuzione. L'utilizzo viene pagato con crediti prepagati; 1 credito equivale a 0,01 USD.
Quali impostazioni supporta Chatterbox?
Secondo il suo schema di input, Chatterbox conosce questi parametri: prompt (fino a 4000 caratteri), seed, audio, cfg_weight (0,2 a 1), temperature (0,05 a 5) e exaggeration (0,25 a 2).
Quanto è veloce Chatterbox?
Non ci sono ancora abbastanza esecuzioni misurate di Chatterbox su Railwail per indicare un tempo di esecuzione. Dipende dall'input, dalle impostazioni e dal carico presso il provider.
Chatterbox è migliore di AudioLDM 2?
Dipende dall'attività . Chatterbox (Resemble AI) e AudioLDM 2 (Haohe Liu) sono entrambi modelli nella categoria Sintesi vocale (TTS). La pagina di confronto mostra i loro prezzi e le specifiche affiancati.
Confronta Chatterbox e AudioLDM 2Modelli comparabili
Tutti in questa categoria- OpenAI TTS-1OpenAI
OpenAI's text-to-speech model. Six built-in voices with natural intonation.
- OpenAI TTS-1 HDOpenAI
OpenAI's high-definition TTS model. Better quality for production use cases.
- Qwen3 TTSAlibaba (Qwen)
A unified Text-to-Speech demo featuring three powerful modes: Voice, Clone and Design
Tutti i modelli tramite un'API
Una chiave API per tutti i modelli su Railwail. L'utilizzo viene addebitato da crediti prepagati, 1 credito = 0,01Â USD.