Qwen3 TTS
qwen3-ttsA unified Text-to-Speech demo featuring three powerful modes: Voice, Clone and Design
- Price
- US$0.024/1k chars
- Input โ output
- Text โ Audio
- Developer
- Alibaba (Qwen)
- Updated
- September 24, 2026
Clone a voice
- 1
Your voice
Voice sample *Record up to 30 s ยท file up to 60 s, 10 MB10 to 30 seconds of clear speech, one speaker, no music or background noise.
Only your own voice or a voice you have verifiable permission for. Imitating other people without their consent is not allowed, see our Terms of Service.
- 2
Text
Language: English103 - 3
Listen
US$0.0025 per run
Playground
Try Qwen3 TTS
Input & output
This run
US$0.0001 ยท 0.01 credits
New here?
10 free credits (US$0.10) when you sign up with Google
Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits. Enough for 1.000 runs of this model.
Examples
- Length: 0:03
Prompt
Hello, I'm Aiden and it's very nice to meet you
Settings
- language
- auto
- mode
- custom_voice
- Length: 0:03
Prompt
Try out Qwen TTS, now on Replicate
Settings
- language
- auto
- mode
- voice_clone
- Length: 0:03
Prompt
Hola, soy Dylan y es un placer conocerte
Settings
- language
- Spanish
- mode
- custom_voice
About Qwen3 TTS
Qwen3 TTS is a model by Alibaba (Qwen) in the Text-to-speech category. On Railwail, Qwen3 TTS costs US$0.024 per 1,000 characters.
Pricing
| 1,000 characters | US$0.024 per 1,000 characters |
|---|
- 1 credit = US$0.01
Cost calculator
Price calculator
Total
US$2.40
240 credits
Per run
US$0.024 ยท 2.4 credits
Fixed price per run, known before the run starts.
API
No verified API example
The public API passes a different input format than this model needs. Use the playground above.
Specifications
- Model ID
qwen3-tts- Developer
- Alibaba (Qwen)
- Category
- Text-to-speech
- Input
- Text
- Output
- Audio
- Billing
- Fixed price, known before the run
- Catalog entry updated
- September 24, 2026
Input parameters
Inputs and settings from the model's input schema. The example in the API section shows which of them the API accepts.
inputrequiredText to synthesize
Type: TextDefault: โAllowed values: โmodecustom_voice: preset speaker; voice_clone: the voice of your reference clip
Type: ChoiceDefault:custom_voiceAllowed values: custom_voice or voice_cloneaudioReference clip to clone the voice from; switches the mode to voice_clone; requires consent
Type: โDefault: โAllowed values: โspeakerOnly with mode = custom_voice
Type: ChoiceDefault:SerenaAllowed values: Aiden, Dylan, Eric, Ono_anna, Ryan, Serena, Sohee, Uncle_fu or VivianlanguageType: ChoiceDefault:autoAllowed values: auto, Chinese, English, Japanese, Korean, French, German, Italian, Spanish, Portuguese or Russianreference_textTranscript of the reference audio (voice_clone, recommended)
Type: TextDefault: โAllowed values: โstyle_instructionType: TextDefault: โAllowed values: โ
Tags
- replicate
- qwen
- text-to-speech
Use cases
Frequently asked questions
What is Qwen3 TTS?
Qwen3 TTS is a model by Alibaba (Qwen) in the Text-to-speech category.
How much does Qwen3 TTS cost on Railwail?
On Railwail, Qwen3 TTS costs US$0.024 per 1,000 characters. The price is known before the run starts. Usage is paid from prepaid credits; 1 credit equals US$0.01.
Which settings does Qwen3 TTS support?
According to its input schema, Qwen3 TTS knows these parameters: input, mode (custom_voice or voice_clone), audio, speaker (Aiden, Dylan, Eric, Ono_anna, Ryan, Serena, Sohee, Uncle_fu or Vivian), language (auto, Chinese, English, Japanese, Korean, French, German, Italian, Spanish, Portuguese or Russian), reference_text and style_instruction.
How fast is Qwen3 TTS?
There are not enough measured runs of Qwen3 TTS on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.
Is Qwen3 TTS better than AudioLDM 2?
That depends on the task. Qwen3 TTS (Alibaba (Qwen)) and AudioLDM 2 (Haohe Liu) are both models in the Text-to-speech category. The comparison page shows their prices and specifications side by side.
Compare Qwen3 TTS and AudioLDM 2Comparable models
All in this category- ChatterboxResemble AI
Resemble AI's open Chatterbox TTS. Zero-shot voice cloning from a short audio prompt with an exaggeration control for emotion intensity, plus CFG weight to balance pacing and fidelity.
- OpenAI TTS-1OpenAI
OpenAI's text-to-speech model. Six built-in voices with natural intonation.
- OpenAI TTS-1 HDOpenAI
OpenAI's high-definition TTS model. Better quality for production use cases.
All models through one API
One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = US$0.01.