Chatterbox

Sprachsyntes (TTS)Tillgänglig
av Resemble AIModell-ID: chatterbox

Resemble AI's open Chatterbox TTS. Zero-shot voice cloning from a short audio prompt with an exaggeration control for emotion intensity, plus CFG weight to balance pacing and fidelity.

Pris
0,030 US$/1 000 tecken
Inmatning → utmatning
Text → Audio
Utvecklare
Resemble AI
Uppdaterad
23 september 2026
01

Clone a voice

Record or upload a short sample, type a text and hear it in that voice. Chatterbox runs in your account like any other run, and you see the price first.
  1. 1

    Your voice

    Voice sample *Record up to 30 s · file up to 60 s, 10 MB

    10 to 30 seconds of clear speech, one speaker, no music or background noise.

  2. 2

    Text

    Example text, change it as you like.103 / 4 000
  3. 3

    Listen

    0,0031 US$ per run

02

Playground

Prova Chatterbox

Inmatning & resultat

0,030 US$/1 000 tecken
Prova Chatterbox

0 / 4 000

Voice sampleRecord up to 30 s · file up to 60 s, 10 MB

10 to 30 seconds of clear speech, one speaker, no music or background noise.

Avancerade inställningar (4)
Resultat
Det genererade talet visas här.

Denna körning

0,0001 US$ · 0,01 credits

Ny här?

10 gratis credits (0,10 US$) när du registrerar dig med Google

Användbar 24 timmar efter registrering, upp till 5 körningar per dag och högst 2 credits per körning. Andra inloggningsmetoder startar utan credits. Räcker för 1 000 körningar av denna modell.

03

Examples

Real outputs from the public examples of this model on Replicate, with the prompt and settings that produced them. They were not generated live on this page.
  • Prompt

    We're excited to introduce Chatterbox, our first production-grade open source TTS model. Licensed under MIT, Chatterbox has been benchmarked against leading closed-source systems like ElevenLabs, and is consistently preferred in side-by-side evaluations. Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. It's also the first open source TTS model to support emotion exaggeration control, a powerful feature that makes your voices stand out. Try it now on our Hugging Face Gradio app. If you like the model but need to scale or finetune it for higher accuracy, check out our competitively priced TTS service (link). It delivers reliable performance with ultra-low latency of sub 200ms—ideal for production use in agents, applications, or interactive media.
  • Prompt

    Now let's make my mum's favourite. So three mars bars into the pan. Then we add the tuna and just stir for a bit, just let the chocolate and fish infuse. A sprinkle of olive oil and some tomato ketchup. Now smell that. Oh boy this is going to be incredible.
04

Om Chatterbox

Kort sagtFrån och med 23 september 2026

Chatterbox är en modell av Resemble AI i kategorin Sprachsyntes (TTS). På Railwail kostar Chatterbox 0,030 US$ per 1 000 tecken.

Chatterbox is Resemble AI's production-grade open TTS. It clones a voice from a brief audio_prompt and offers an exaggeration slider that scales emotional intensity, which is unusual for open models. cfg_weight and temperature tune delivery. Hosted by resemble-ai on Replicate.
05

Priser

Priser i US-dollar. Användningen debiteras från förbetald kredit.
1 000 tecken0,030 US$ per 1 000 tecken
  • 1 kredit = 0,01 US$

Kostnadsräknare

Prisräknare

/ körning

Totalt

3,00 US$

300 krediter

Per körning

0,03 US$ · 3 krediter

Fast pris per körning, känt innan körningen startar.

06

API

Anropa Chatterbox med din Railwail API-nyckel. Använd detta modell-ID i begäran:

Inget verifierat API-exempel

Det offentliga API:et skickar ett annat inmatningsformat än vad denna modell behöver. Använd lekplatsen ovan.

07

Specifikationer

Modell-ID
chatterbox
Utvecklare
Resemble AI
Inmatning
Text
Utmatning
Audio
Fakturering
Fastpris, känt innan körningen
Kataloginlägg uppdaterat
23 september 2026

Indataparametrar

Inmatningar och inställningar från modellens indataschema. Exemplet i API-avsnittet visar vilka som API:et accepterar.

  • promptobligatorisk

    Text to speak

    Typ: Text
    Standard: –
    Tillåtna värden: upp till 4 000 tecken
  • seed

    0 = random

    Typ: Heltal
    Standard: 0
    Tillåtna värden: –
  • audio

    Optional reference clip (5-60 s) to clone the voice from; requires consent

    Typ: –
    Standard: –
    Tillåtna värden: –
  • cfg_weight
    Typ: Tal
    Standard: 0.5
    Tillåtna värden: 0,2 till 1
  • temperature
    Typ: Tal
    Standard: 0.8
    Tillåtna värden: 0,05 till 5
  • exaggeration
    Typ: Tal
    Standard: 0.5
    Tillåtna värden: 0,25 till 2

Taggar

  • replicate
  • resemble-ai
  • tts
  • voice-cloning
  • expressive
08

Användningsfall

09

Vanliga frågor

Vad är Chatterbox?

Chatterbox är en modell av Resemble AI i kategorin Sprachsyntes (TTS).

Vad kostar Chatterbox på Railwail?

På Railwail kostar Chatterbox 0,030 US$ per 1 000 tecken. Priset är känt innan körningen startar. Användningen betalas från förbetald kredit; 1 kredit motsvarar 0,01 US$.

Vilka inställningar stöder Chatterbox?

Enligt sitt inmatningsschema känner Chatterbox till dessa parametrar: prompt (upp till 4 000 tecken), seed, audio, cfg_weight (0,2 till 1), temperature (0,05 till 5) och exaggeration (0,25 till 2).

Hur snabb är Chatterbox?

Det finns ännu inte tillräckligt många uppmätta körningar av Chatterbox på Railwail för att ange en körningstid. Det beror på inmatningen, inställningarna och belastningen hos leverantören.

Är Chatterbox bättre än AudioLDM 2?

Det beror på uppgiften. Chatterbox (Resemble AI) och AudioLDM 2 (Haohe Liu) är båda modeller i kategorin Sprachsyntes (TTS). Jämförelsesidan visar deras priser och specifikationer sida vid sida.

Jämför Chatterbox och AudioLDM 2
10

Jämförbara modeller

Alla i denna kategori

Alla modeller via ett API

En API-nyckel för alla modeller på Railwail. Användningen debiteras från förbetald kredit, 1 kredit = 0,01 US$.