Chatterbox

Sintesi vocale (TTS)Disponibile
di Resemble AIID modello: chatterbox

Resemble AI's open Chatterbox TTS. Zero-shot voice cloning from a short audio prompt with an exaggeration control for emotion intensity, plus CFG weight to balance pacing and fidelity.

Prezzo
0,030 USD/1.000 caratteri
Input → output
Testo → Audio
Sviluppatore
Resemble AI
Aggiornato
23 settembre 2026
01

Clone a voice

Record or upload a short sample, type a text and hear it in that voice. Chatterbox runs in your account like any other run, and you see the price first.
  1. 1

    Your voice

    Voice sample *Record up to 30 s · file up to 60 s, 10 MB

    10 to 30 seconds of clear speech, one speaker, no music or background noise.

  2. 2

    Text

    Example text, change it as you like.103 / 4000
  3. 3

    Listen

    0,0031 USD per run

02

Playground

Prova Chatterbox

Input e output

0,030 USD/1.000 caratteri
Prova Chatterbox

0 / 4000

Voice sampleRecord up to 30 s · file up to 60 s, 10 MB

10 to 30 seconds of clear speech, one speaker, no music or background noise.

Impostazioni avanzate (4)
Risultato
La voce generata appare qui.

Questa esecuzione

0,0001 USD · 0,01 crediti

Nuovo qui?

10 crediti gratuiti (0,10 USD) quando ti iscrivi con Google

Utilizzabile 24 ore dopo l'iscrizione, fino a 5 esecuzioni al giorno e al massimo 2 crediti per esecuzione. Altri metodi di accesso iniziano senza crediti. Sufficiente per 1000 esecuzioni di questo modello.

03

Examples

Real outputs from the public examples of this model on Replicate, with the prompt and settings that produced them. They were not generated live on this page.
  • Prompt

    We're excited to introduce Chatterbox, our first production-grade open source TTS model. Licensed under MIT, Chatterbox has been benchmarked against leading closed-source systems like ElevenLabs, and is consistently preferred in side-by-side evaluations. Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. It's also the first open source TTS model to support emotion exaggeration control, a powerful feature that makes your voices stand out. Try it now on our Hugging Face Gradio app. If you like the model but need to scale or finetune it for higher accuracy, check out our competitively priced TTS service (link). It delivers reliable performance with ultra-low latency of sub 200ms—ideal for production use in agents, applications, or interactive media.
  • Prompt

    Now let's make my mum's favourite. So three mars bars into the pan. Then we add the tuna and just stir for a bit, just let the chocolate and fish infuse. A sprinkle of olive oil and some tomato ketchup. Now smell that. Oh boy this is going to be incredible.
04

Informazioni su Chatterbox

RiassuntoA partire da 23 settembre 2026

Chatterbox è un modello di Resemble AI nella categoria Sintesi vocale (TTS). Su Railwail, Chatterbox costa 0,030 USD per 1.000 caratteri.

Chatterbox is Resemble AI's production-grade open TTS. It clones a voice from a brief audio_prompt and offers an exaggeration slider that scales emotional intensity, which is unusual for open models. cfg_weight and temperature tune delivery. Hosted by resemble-ai on Replicate.
05

Prezzi

Prezzi in dollari USA. L'utilizzo viene addebitato dai crediti prepagati.
1.000 caratteri0,030 USD per 1.000 caratteri
  • 1 credito = 0,01 USD

Calcolatore di costi

Calcolatore prezzi

/ esecuzione

Totale

3,00 USD

300 crediti

Per esecuzione

0,03 USD · 3 crediti

Prezzo fisso per esecuzione, noto prima dell'inizio.

06

API

Chiama Chatterbox con la tua chiave API Railwail. Usa questo ID modello nella richiesta:

Nessun esempio API verificato

L'API pubblica passa un formato di input diverso da quello richiesto da questo modello. Usa il playground sopra.

07

Specifiche

ID modello
chatterbox
Sviluppatore
Resemble AI
Input
Testo
Output
Audio
Fatturazione
Prezzo fisso, noto prima dell'esecuzione
Voce di catalogo aggiornata
23 settembre 2026

Parametri di input

Input e impostazioni dallo schema di input del modello. L'esempio nella sezione API mostra quali di essi l'API accetta.

  • promptObbligatorio

    Text to speak

    Tipo: Testo
    Predefinito: –
    Valori consentiti: fino a 4000 caratteri
  • seed

    0 = random

    Tipo: Numero intero
    Predefinito: 0
    Valori consentiti: –
  • audio

    Optional reference clip (5-60 s) to clone the voice from; requires consent

    Tipo: –
    Predefinito: –
    Valori consentiti: –
  • cfg_weight
    Tipo: Numero
    Predefinito: 0.5
    Valori consentiti: 0,2 a 1
  • temperature
    Tipo: Numero
    Predefinito: 0.8
    Valori consentiti: 0,05 a 5
  • exaggeration
    Tipo: Numero
    Predefinito: 0.5
    Valori consentiti: 0,25 a 2

Etichette

  • replicate
  • resemble-ai
  • tts
  • voice-cloning
  • expressive
08

Casi d'uso

09

Domande frequenti

Cos'è Chatterbox?

Chatterbox è un modello di Resemble AI nella categoria Sintesi vocale (TTS).

Quanto costa Chatterbox su Railwail?

Su Railwail, Chatterbox costa 0,030 USD per 1.000 caratteri. Il prezzo è noto prima dell'inizio dell'esecuzione. L'utilizzo viene pagato con crediti prepagati; 1 credito equivale a 0,01 USD.

Quali impostazioni supporta Chatterbox?

Secondo il suo schema di input, Chatterbox conosce questi parametri: prompt (fino a 4000 caratteri), seed, audio, cfg_weight (0,2 a 1), temperature (0,05 a 5) e exaggeration (0,25 a 2).

Quanto è veloce Chatterbox?

Non ci sono ancora abbastanza esecuzioni misurate di Chatterbox su Railwail per indicare un tempo di esecuzione. Dipende dall'input, dalle impostazioni e dal carico presso il provider.

Chatterbox è migliore di AudioLDM 2?

Dipende dall'attività. Chatterbox (Resemble AI) e AudioLDM 2 (Haohe Liu) sono entrambi modelli nella categoria Sintesi vocale (TTS). La pagina di confronto mostra i loro prezzi e le specifiche affiancati.

Confronta Chatterbox e AudioLDM 2
10

Modelli comparabili

Tutti in questa categoria

Tutti i modelli tramite un'API

Una chiave API per tutti i modelli su Railwail. L'utilizzo viene addebitato da crediti prepagati, 1 credito = 0,01 USD.