ElevenLabs v3 (alpha)

Sinteză vocală (TTS)Indisponibil
de ElevenLabsID model: eleven-v3

ElevenLabs' v3 alpha TTS. Most expressive voice model with audio tags and laughter, higher latency.

Status
Indisponibil
Intrare → ieșire
Text → Audio
Dezvoltator
ElevenLabs
Actualizat
25 iunie 2026

ElevenLabs v3 (alpha) nu este disponibil în acest moment

Indisponibil în prezent: acest model a fost dezactivat.

Poți citi în continuare detaliile pe această pagină. Alege una dintre alternativele disponibile de mai jos pentru a rula imediat un model comparabil.

Mergi la alternative
01
  • AudioLDM 2AudioLDM

    Latent-diffusion model for general-purpose text-to-audio. Generates speech, music, and sound effects with a unified prior.

    ≈ 0,0157 USD/rulare

  • ChatterboxReplicate

    Resemble AI's open Chatterbox TTS. Zero-shot voice cloning from a short audio prompt with an exaggeration control for emotion intensity, plus CFG weight to balance pacing and fidelity.

    0,030 USD/1.000 caractere

  • F5-TTSReplicate

    Open-source flow-matching TTS with strong zero-shot voice cloning. Code MIT, weights CC-BY-NC.

    ≈ 0,0168 USD/rulare

02

Playground

Încearcă ElevenLabs v3 (alpha)

Intrare & rezultat

Indisponibil în prezent

Indisponibil în prezent: acest model a fost dezactivat.

Playground-ul este dezactivat. Modele comparabile găsești în aceeași categorie: Vezi alternativele

Încearcă ElevenLabs v3 (alpha)

0 / 4.000

Rezultat
Vocea generată apare aici.

Această rulare

Fără preț – momentan indisponibil.

Nou aici?

5 credite gratuite (0,05 USD) când te înregistrezi cu Google

Utilizabil 24 ore după înregistrare, până la 5 rulări pe zi și maximum 2 credite pe rulare. Alte metode de conectare încep fără credite.

03

Despre ElevenLabs v3 (alpha)

Pe scurtDin 25 iunie 2026

ElevenLabs v3 (alpha) este un model de ElevenLabs din categoria Sinteză vocală (TTS). ElevenLabs v3 (alpha) nu este disponibil în prezent pe Railwail.

Fundal

Despre ElevenLabs

Fondat 2022 · London, UK / New York, USA

ElevenLabs was founded in 2022 by Piotr Dabkowski (CTO, ex-Google ML engineer) and Mati Staniszewski (CEO, ex-Palantir), two Polish high-school friends frustrated with the poor quality of TV-show dubbing in Polish. The company set out to build voice AI that captures intonation and emotion across languages. Headquartered in London and New York with engineering hubs in Warsaw and the Bay Area, ElevenLabs raised a $19M Series A in June 2023 led by Andreessen Horowitz, a $80M Series B in January 2024 also led by a16z at a $1.1B valuation, and a $180M Series C in January 2025 at a $3.3B valuation co-led by a16z and ICONIQ. ElevenLabs v3 (alpha) was previewed in 2025 as the next generation flagship model with expressive emotion tags, longer context and more languages, succeeding the Multilingual V2 family that became the de-facto standard for AI dubbing.

Vizitează ElevenLabs

Arhitectură

Proprietary autoregressive Transformer TTS with neural codec and emotion/prosody conditioning

ElevenLabs v3 (alpha) is the company's 2025 flagship text-to-speech model and the first ElevenLabs system to expose explicit emotion and event tags inside text input ([whispers], [laughs], [angry], [sighs]). It is a proprietary Transformer-based autoregressive model that predicts neural-codec audio tokens conditioned on a text prompt and a speaker embedding obtained from a few seconds of reference audio (Instant Voice Clone) or a fully fine-tuned voice (Professional Voice Clone, requires ~30 minutes of clean audio). v3 expands language coverage from 29 (v2) to 70+ languages, lengthens the input window to roughly 10,000 characters per request, and adds dialogue mode for multi-speaker scenes. ElevenLabs has not published a technical paper; product blog posts describe internal improvements in speaker disentanglement, code-switching and emotional range. v3 is offered through the same hosted API and Studio UI as Multilingual V2 but at higher latency and price.

Parametri
Undisclosed
Context
10.000 tokeni

Capabilități

  • Expressive emotion and event tags ([laughs], [whispers], [angry], [crying])
  • 70+ languages with high-quality code-switching
  • Multi-speaker dialogue mode for podcast and audiobook generation
  • Instant Voice Clone from ~1 minute of audio and Professional Voice Clone from ~30 minutes
  • Long-form input up to ~10,000 characters per request
  • Studio editor for multi-paragraph projects with per-line speaker control
  • Best for: audiobooks, dubbing, narrative podcasts, character voices for games

Antrenament & licență

Not disclosed. ElevenLabs licences professional voice talent, uses public-domain audiobooks and crowd-sourced opt-in voice contributions; commercial recordings are excluded per their public statements.

Licență: Proprietary commercial SaaS. Commercial use of generated audio is permitted on paid plans; voice clones remain customer property.

Teste de siguranță: AI Speech Classifier offered for free; mandatory consent statement and KYC for Professional Voice Clones; inaudible watermark on outputs; mis-use bans after 2024 deepfake incidents.

Limitări cunoscute

  • Higher latency than v2 Turbo or Cartesia Sonic
  • Tag interpretation occasionally inconsistent in alpha
  • Hard refusal for likeness of named public figures without verified consent
  • Closed weights, no on-premise deployment
  • Pricing per character is among the highest in the market
04

Prețuri

Indisponibil în prezent: acest model a fost dezactivat. Nu există preț pentru acest model în acest moment, deci nu poate fi executat.

05

API

Apelează ElevenLabs v3 (alpha) cu cheia ta API Railwail. Folosește acest ID de model în cerere:

Niciun exemplu API verificat

API-ul public transmite un format de intrare diferit de ceea ce are nevoie acest model. Folosește playground-ul de mai sus.

06

Specificații

ID model
eleven-v3
Dezvoltator
ElevenLabs
Intrare
Text
Ieșire
Audio
Formate de ieșire
MP3, WAV, OPUS
Dimensiune model
Undisclosed
Licență
Proprietary commercial SaaS. Commercial use of generated audio is permitted on paid plans; voice clones remain customer property.
Intrare catalog actualizată
25 iunie 2026

Parametri de intrare

Intrări și setări din schema de intrare a modelului. Exemplul din secțiunea API arată care dintre ele acceptă API-ul.

  • inputobligatoriu

    Text to convert to speech

    Tip: Text
    Implicit:
    Valori permise: până la 4.000 caractere
  • speed
    Tip: Număr
    Implicit: 1
    Valori permise: 0,5 până la 2
  • voice
    Tip: Alegere
    Implicit: Rachel
    Valori permise: Rachel, Adam, Antoni, Bella, Domi, Elli, Josh sau Sam
  • response_format
    Tip: Alegere
    Implicit: mp3
    Valori permise: mp3, wav sau opus

Etichete

  • elevenlabs
  • tts
  • expressive
  • alpha
  • per-character
07

Cazuri de utilizare

Pentru ce se folosește

  • Audiobook narration with emotion
  • Multilingual film and series dubbing
  • Character voices for video games
  • Narrative podcasts and radio drama
  • Accessibility tools and screen readers
08

Întrebări frecvente

Ce este ElevenLabs v3 (alpha)?

ElevenLabs v3 (alpha) este un model de ElevenLabs din categoria Sinteză vocală (TTS). Este listat pe Railwail, dar nu poate fi rulat în acest moment.

Cât costă ElevenLabs v3 (alpha) pe Railwail?

ElevenLabs v3 (alpha) nu poate fi rulat pe Railwail în acest moment, deci nu există preț curent. Alternativele disponibile cu prețuri sunt listate mai jos pe această pagină.

Ce setări acceptă ElevenLabs v3 (alpha)?

Conform schemei sale de intrare, ElevenLabs v3 (alpha) cunoaște acești parametri: input (până la 4.000 caractere), speed (0,5 până la 2), voice (Rachel, Adam, Antoni, Bella, Domi, Elli, Josh sau Sam) și response_format (mp3, wav sau opus).

Cât de rapid este ElevenLabs v3 (alpha)?

Nu sunt suficiente rulări măsurate ale ElevenLabs v3 (alpha) pe Railwail încă pentru a indica un timp de rulare. Depinde de intrare, de setări și de sarcina la furnizor.

Este ElevenLabs v3 (alpha) mai bun decât AudioLDM 2?

Depinde de sarcină. ElevenLabs v3 (alpha) (ElevenLabs) și AudioLDM 2 (AudioLDM) sunt ambele modele din categoria Sinteză vocală (TTS). Pagina de comparație arată prețurile și specificațiile lor una lângă alta.

Compară ElevenLabs v3 (alpha) și AudioLDM 2

Pot folosi ElevenLabs v3 (alpha) chiar acum?

Indisponibil în prezent: acest model a fost dezactivat. Pagina rămâne online; alternativele disponibile din aceeași categorie sunt listate mai jos.

Toate modelele printr-o singură API

O cheie API pentru fiecare model pe Railwail. Utilizarea se percepe din credite prepay, 1 credit = 0,01 USD.