ElevenLabs v3 (alpha)

Převod textu na řečNedostupné
od ElevenLabsID modelu: eleven-v3

ElevenLabs' v3 alpha TTS. Most expressive voice model with audio tags and laughter, higher latency.

Stav
Nedostupné
Vstup → výstup
Text → Audio
Vývojář
ElevenLabs
Aktualizováno
25. června 2026

ElevenLabs v3 (alpha) není momentálně dostupný

Momentálně nedostupné: tento model byl deaktivován.

Podrobnosti na této stránce si můžete přečíst. Vyberte jednu z dostupných alternativ níže a spusťte porovnatelný model hned.

Na alternativy
01

Srovnatelné modely

Všechny v této kategorii
  • AudioLDM 2AudioLDM

    Latent-diffusion model for general-purpose text-to-audio. Generates speech, music, and sound effects with a unified prior.

    ≈ 0,0157 US$/spuštění

  • ChatterboxReplicate

    Resemble AI's open Chatterbox TTS. Zero-shot voice cloning from a short audio prompt with an exaggeration control for emotion intensity, plus CFG weight to balance pacing and fidelity.

    0,030 US$/1 000 znaků

  • F5-TTSReplicate

    Open-source flow-matching TTS with strong zero-shot voice cloning. Code MIT, weights CC-BY-NC.

    ≈ 0,0168 US$/spuštění

02

Playground

Vyzkoušet ElevenLabs v3 (alpha)

Vstup a výstup

Momentálně nedostupné

Momentálně nedostupné: tento model byl deaktivován.

Playground je zakázán. Srovnatelné modely najdete ve stejné kategorii: Zobrazit alternativy

Vyzkoušet ElevenLabs v3 (alpha)

0 / 4 000

Výsledek
Vygenerovaná řeč se zobrazí zde.

Tento běh

Bez ceny – momentálně nedostupné.

Nový zde?

5 bezplatných credits (0,05 US$) při registraci přes Google

Použitelné 24 hodin po registraci, až 5 běhů za den a maximálně 2 credits za běh. Ostatní způsoby přihlášení začínají bez credits.

03

O ElevenLabs v3 (alpha)

StručněStav: 25. června 2026

ElevenLabs v3 (alpha) je model od ElevenLabs v kategorii Převod textu na řeč. ElevenLabs v3 (alpha) není na Railwail v současné době dostupný.

Pozadí

O ElevenLabs

Založeno 2022 · London, UK / New York, USA

ElevenLabs was founded in 2022 by Piotr Dabkowski (CTO, ex-Google ML engineer) and Mati Staniszewski (CEO, ex-Palantir), two Polish high-school friends frustrated with the poor quality of TV-show dubbing in Polish. The company set out to build voice AI that captures intonation and emotion across languages. Headquartered in London and New York with engineering hubs in Warsaw and the Bay Area, ElevenLabs raised a $19M Series A in June 2023 led by Andreessen Horowitz, a $80M Series B in January 2024 also led by a16z at a $1.1B valuation, and a $180M Series C in January 2025 at a $3.3B valuation co-led by a16z and ICONIQ. ElevenLabs v3 (alpha) was previewed in 2025 as the next generation flagship model with expressive emotion tags, longer context and more languages, succeeding the Multilingual V2 family that became the de-facto standard for AI dubbing.

Navštívit ElevenLabs

Architektura

Proprietary autoregressive Transformer TTS with neural codec and emotion/prosody conditioning

ElevenLabs v3 (alpha) is the company's 2025 flagship text-to-speech model and the first ElevenLabs system to expose explicit emotion and event tags inside text input ([whispers], [laughs], [angry], [sighs]). It is a proprietary Transformer-based autoregressive model that predicts neural-codec audio tokens conditioned on a text prompt and a speaker embedding obtained from a few seconds of reference audio (Instant Voice Clone) or a fully fine-tuned voice (Professional Voice Clone, requires ~30 minutes of clean audio). v3 expands language coverage from 29 (v2) to 70+ languages, lengthens the input window to roughly 10,000 characters per request, and adds dialogue mode for multi-speaker scenes. ElevenLabs has not published a technical paper; product blog posts describe internal improvements in speaker disentanglement, code-switching and emotional range. v3 is offered through the same hosted API and Studio UI as Multilingual V2 but at higher latency and price.

Parametry
Undisclosed
Kontext
10,000 tokenů

Schopnosti

  • Expressive emotion and event tags ([laughs], [whispers], [angry], [crying])
  • 70+ languages with high-quality code-switching
  • Multi-speaker dialogue mode for podcast and audiobook generation
  • Instant Voice Clone from ~1 minute of audio and Professional Voice Clone from ~30 minutes
  • Long-form input up to ~10,000 characters per request
  • Studio editor for multi-paragraph projects with per-line speaker control
  • Best for: audiobooks, dubbing, narrative podcasts, character voices for games

Trénování a licence

Not disclosed. ElevenLabs licences professional voice talent, uses public-domain audiobooks and crowd-sourced opt-in voice contributions; commercial recordings are excluded per their public statements.

Licence: Proprietary commercial SaaS. Commercial use of generated audio is permitted on paid plans; voice clones remain customer property.

Bezpečnostní testy: AI Speech Classifier offered for free; mandatory consent statement and KYC for Professional Voice Clones; inaudible watermark on outputs; mis-use bans after 2024 deepfake incidents.

Známá omezení

  • Higher latency than v2 Turbo or Cartesia Sonic
  • Tag interpretation occasionally inconsistent in alpha
  • Hard refusal for likeness of named public figures without verified consent
  • Closed weights, no on-premise deployment
  • Pricing per character is among the highest in the market
04

Ceny

Momentálně nedostupné: tento model byl deaktivován. Pro tento model momentálně není cena, takže jej nelze spustit.

05

API

Volejte ElevenLabs v3 (alpha) s vaším API klíčem Railwail. V požadavku použijte toto ID modelu:

Žádný ověřený příklad API

Veřejné API předává jiný formát vstupu, než tento model potřebuje. Použijte playground výše.

06

Specifikace

ID modelu
eleven-v3
Vývojář
ElevenLabs
Vstup
Text
Výstup
Audio
Formáty výstupu
MP3, WAV, OPUS
Velikost modelu
Undisclosed
Licence
Proprietary commercial SaaS. Commercial use of generated audio is permitted on paid plans; voice clones remain customer property.
Záznam v katalogu aktualizován
25. června 2026

Vstupní parametry

Vstupy a nastavení ze vstupního schématu modelu. Příklad v sekci API ukazuje, které z nich API přijímá.

  • inputpovinné

    Text to convert to speech

    Typ: Text
    Výchozí:
    Povolené hodnoty: až 4,000 znaků
  • speed
    Typ: Číslo
    Výchozí: 1
    Povolené hodnoty: 0.5 až 2
  • voice
    Typ: Volba
    Výchozí: Rachel
    Povolené hodnoty: Rachel, Adam, Antoni, Bella, Domi, Elli, Josh nebo Sam
  • response_format
    Typ: Volba
    Výchozí: mp3
    Povolené hodnoty: mp3, wav nebo opus

Štítky

  • elevenlabs
  • tts
  • expressive
  • alpha
  • per-character
07

Případy použití

K čemu se používá

  • Audiobook narration with emotion
  • Multilingual film and series dubbing
  • Character voices for video games
  • Narrative podcasts and radio drama
  • Accessibility tools and screen readers
08

Často kladené otázky

Co je ElevenLabs v3 (alpha)?

ElevenLabs v3 (alpha) je model od ElevenLabs v kategorii Převod textu na řeč. Je uveden na Railwail, ale momentálně jej nelze spustit.

Kolik stojí ElevenLabs v3 (alpha) na Railwail?

ElevenLabs v3 (alpha) se momentálně na Railwail nedá spustit, takže není aktuální cena. Dostupné alternativy s cenami jsou uvedeny dále na této stránce.

Jaká nastavení ElevenLabs v3 (alpha) podporuje?

Podle schématu vstupu ElevenLabs v3 (alpha) zná tyto parametry: input (až 4,000 znaků), speed (0.5 až 2), voice (Rachel, Adam, Antoni, Bella, Domi, Elli, Josh nebo Sam) a response_format (mp3, wav nebo opus).

Jak rychlý je ElevenLabs v3 (alpha)?

Pro ElevenLabs v3 (alpha) je na Railwail zatím příliš málo naměřených spuštění na uvedení doby běhu. Závisí na vstupu, nastavení a zátěži u poskytovatele.

Je ElevenLabs v3 (alpha) lepší než AudioLDM 2?

Záleží na úkolu. ElevenLabs v3 (alpha) (ElevenLabs) a AudioLDM 2 (AudioLDM) jsou oba modely v kategorii Převod textu na řeč. Stránka porovnání zobrazuje jejich ceny a specifikace vedle sebe.

Porovnat ElevenLabs v3 (alpha) a AudioLDM 2

Mohu ElevenLabs v3 (alpha) používat hned teď?

Momentálně nedostupné: tento model byl deaktivován. Stránka zůstává online; dostupné alternativy ze stejné kategorie jsou uvedeny dále.

Všechny modely přes jednu API

Jeden API klíč pro všechny modely na Railwail. Použití se účtuje z předplacených kreditů, 1 kredit = 0,01 US$.