XTTS v2

Text-to-speechUnavailable
by CommunityModel ID: xtts-v2

Coqui's XTTS v2 multilingual TTS with voice cloning from 6 seconds of reference audio. Supports 17 languages and emotion transfer.

Status
Unavailable
Input β†’ output
Text β†’ Audio
Developer
Community
Updated
September 23, 2026

XTTS v2 is currently unavailable

Currently unavailable: this model has been deactivated.

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

Go to alternatives
01

Comparable models

All in this category
  • AudioLDM 2Haohe Liu

    Latent-diffusion model for general-purpose text-to-audio. Generates speech, music, and sound effects with a unified prior.

    β‰ˆ $0.0157/run

  • Kokoro TTS 82MCommunity

    Open-weights 82M-parameter TTS. Punches above its size class on naturalness benchmarks at a fraction of the inference cost of larger models.

    β‰ˆ $0.00030/run

  • OpenVoice v2Community

    MyShell OpenVoice v2. Multilingual zero-shot voice cloning with accurate tone-color reproduction and style/emotion control.

    β‰ˆ $0.0673/run

02

Playground

Try XTTS v2

Input & output

Currently unavailable

Currently unavailable: this model has been deactivated.

The playground is disabled. You can find comparable models in the same category: Browse alternatives

Try XTTS v2

0 / 4,000

Text to synthesize

Voice sample *Record up to 30 s Β· file up to 60 s, 10 MB

10 to 30 seconds of clear speech, one speaker, no music or background noise.

Advanced settings (1)
Output
The generated speech appears here.

This run

No price – currently unavailable.

New here?

10 free credits ($0.10) when you sign up with Google

Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits.

03

Examples

Real outputs from the public examples of this model on Replicate, with the prompt and settings that produced them. They were not generated live on this page.
04

About XTTS v2

TL;DRAs of September 23, 2026

XTTS v2 is a model by Community in the Text-to-speech category. XTTS v2 is currently not available on Railwail.

05

Pricing

Currently unavailable: this model has been deactivated. There is no price for this model at the moment, so it cannot be run.

06

API

Call XTTS v2 with your Railwail API key. Use this model ID in the request:

No verified API example

The public API passes a different input format than this model needs. Use the playground above.

07

Specifications

Model ID
xtts-v2
Developer
Community
Input
Text
Output
Audio
Catalog entry updated
September 23, 2026

Input parameters

Inputs and settings from the model's input schema. The example in the API section shows which of them the API accepts.

  • textrequired

    Text to synthesize

    Type: Text
    Default: –
    Allowed values: up to 4,000 characters
  • audiorequired

    Speaker sample (6-60 s); requires consent

    Type: –
    Default: –
    Allowed values: –
  • language
    Type: Choice
    Default: en
    Allowed values: en, es, fr, de, it, pt, pl, tr, ru, nl, cs, ar, zh, hu, ko, or hi
  • cleanup_voice
    Type: Yes/no
    Default: false
    Allowed values: –

Tags

  • coqui
  • tts
  • voice-cloning
  • multilingual
  • open-weights
  • anon-free
08

Use cases

09

Frequently asked questions

What is XTTS v2?

XTTS v2 is a model by Community in the Text-to-speech category. It is listed on Railwail but cannot be run at the moment.

How much does XTTS v2 cost on Railwail?

XTTS v2 cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

Which settings does XTTS v2 support?

According to its input schema, XTTS v2 knows these parameters: text (up to 4,000 characters), audio, language (en, es, fr, de, it, pt, pl, tr, ru, nl, cs, ar, zh, hu, ko, or hi), and cleanup_voice.

How fast is XTTS v2?

There are not enough measured runs of XTTS v2 on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

Is XTTS v2 better than AudioLDM 2?

That depends on the task. XTTS v2 (Community) and AudioLDM 2 (Haohe Liu) are both models in the Text-to-speech category. The comparison page shows their prices and specifications side by side.

Compare XTTS v2 and AudioLDM 2

Can I use XTTS v2 right now?

Currently unavailable: this model has been deactivated. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.