AI Voice Cloning Models

Voice cloning models listen to a short reference clip and then read any text in that voice. Record the sample in your browser or upload a file (MP3, WAV, WebM or OGG, up to 60 seconds), type your text and listen to the result.

For every voice sample you confirm that you may use the voice: your own, or one you have permission for. Runs are billed like any other model run, from a prepaid balance.

3 models for this use case

3 modeller

  • Chatterbox

    Tekst-til-taleCommunity

    Resemble AI's open Chatterbox TTS. Zero-shot voice cloning from a short audio prompt with an exaggeration control for emotion intensity, plus CFG weight to balance pacing and fidelity.

    $0.030 / kjøring
    replicateresemble-aitts
  • OpenVoice v2

    Tekst-til-taleCommunity

    MyShell OpenVoice v2. Multilingual zero-shot voice cloning with accurate tone-color reproduction and style/emotion control.

    $0.067 / kjøring
    myshellttsvoice-cloning
  • Qwen3 TTS

    Tekst-til-taleCommunity
    Ny

    A unified Text-to-Speech demo featuring three powerful modes: Voice, Clone and Design

    $0.024 / kjøring
    replicateqwentext-to-speech

Frequently asked questions

One API, pay only for what you use

Try any model with a free generation, no signup. Then transparent per-use pricing, no subscription.

Related use cases

AI Voice Cloning - Clone a Voice from a Short Sample | Railwail