ElevenLabs v3 (alpha)

Tekst-naar-spraakNiet beschikbaar
van ElevenLabsModel-ID: eleven-v3

ElevenLabs' v3 alpha TTS. Most expressive voice model with audio tags and laughter, higher latency.

Status
Niet beschikbaar
Invoer → Uitvoer
Tekst → Audio
Ontwikkelaar
ElevenLabs
Bijgewerkt
25 juni 2026

ElevenLabs v3 (alpha) is momenteel niet beschikbaar

Momenteel niet beschikbaar: dit model is gedeactiveerd.

Je kunt de details op deze pagina nog steeds lezen. Kies een van de beschikbare alternatieven hieronder om direct een vergelijkbaar model uit te voeren.

Naar alternatieven
01

Vergelijkbare modellen

Alle in deze categorie
  • AudioLDM 2AudioLDM

    Latent-diffusion model for general-purpose text-to-audio. Generates speech, music, and sound effects with a unified prior.

    ≈ US$ 0,0157/uitvoering

  • ChatterboxReplicate

    Resemble AI's open Chatterbox TTS. Zero-shot voice cloning from a short audio prompt with an exaggeration control for emotion intensity, plus CFG weight to balance pacing and fidelity.

    US$ 0,030/1.000 tekens

  • F5-TTSReplicate

    Open-source flow-matching TTS with strong zero-shot voice cloning. Code MIT, weights CC-BY-NC.

    ≈ US$ 0,0168/uitvoering

02

Playground

ElevenLabs v3 (alpha) proberen

Invoer & resultaat

Momenteel niet beschikbaar

Momenteel niet beschikbaar: dit model is gedeactiveerd.

De playground is uitgeschakeld. Vergelijkbare modellen vind je in dezelfde categorie: Alternatieven bekijken

ElevenLabs v3 (alpha) proberen

0 / 4.000

Resultaat
De gegenereerde spraak verschijnt hier.

Deze uitvoering

Geen prijs – momenteel niet beschikbaar.

Nieuw hier?

5 gratis credits (US$ 0,05) wanneer je je aanmeldt met Google

Bruikbaar 24 uur na aanmelding, tot 5 uitvoeringen per dag en maximaal 2 credits per uitvoering. Andere aanmeldmethoden starten zonder credits.

03

Over ElevenLabs v3 (alpha)

SamengevatPer 25 juni 2026

ElevenLabs v3 (alpha) is een model van ElevenLabs in de categorie Tekst-naar-spraak. ElevenLabs v3 (alpha) is momenteel niet beschikbaar op Railwail.

Achtergrond

Over ElevenLabs

Opgericht 2022 · London, UK / New York, USA

ElevenLabs was founded in 2022 by Piotr Dabkowski (CTO, ex-Google ML engineer) and Mati Staniszewski (CEO, ex-Palantir), two Polish high-school friends frustrated with the poor quality of TV-show dubbing in Polish. The company set out to build voice AI that captures intonation and emotion across languages. Headquartered in London and New York with engineering hubs in Warsaw and the Bay Area, ElevenLabs raised a $19M Series A in June 2023 led by Andreessen Horowitz, a $80M Series B in January 2024 also led by a16z at a $1.1B valuation, and a $180M Series C in January 2025 at a $3.3B valuation co-led by a16z and ICONIQ. ElevenLabs v3 (alpha) was previewed in 2025 as the next generation flagship model with expressive emotion tags, longer context and more languages, succeeding the Multilingual V2 family that became the de-facto standard for AI dubbing.

ElevenLabs bezoeken

Architectuur

Proprietary autoregressive Transformer TTS with neural codec and emotion/prosody conditioning

ElevenLabs v3 (alpha) is the company's 2025 flagship text-to-speech model and the first ElevenLabs system to expose explicit emotion and event tags inside text input ([whispers], [laughs], [angry], [sighs]). It is a proprietary Transformer-based autoregressive model that predicts neural-codec audio tokens conditioned on a text prompt and a speaker embedding obtained from a few seconds of reference audio (Instant Voice Clone) or a fully fine-tuned voice (Professional Voice Clone, requires ~30 minutes of clean audio). v3 expands language coverage from 29 (v2) to 70+ languages, lengthens the input window to roughly 10,000 characters per request, and adds dialogue mode for multi-speaker scenes. ElevenLabs has not published a technical paper; product blog posts describe internal improvements in speaker disentanglement, code-switching and emotional range. v3 is offered through the same hosted API and Studio UI as Multilingual V2 but at higher latency and price.

Parameters
Undisclosed
Context
10.000 tokens

Mogelijkheden

  • Expressive emotion and event tags ([laughs], [whispers], [angry], [crying])
  • 70+ languages with high-quality code-switching
  • Multi-speaker dialogue mode for podcast and audiobook generation
  • Instant Voice Clone from ~1 minute of audio and Professional Voice Clone from ~30 minutes
  • Long-form input up to ~10,000 characters per request
  • Studio editor for multi-paragraph projects with per-line speaker control
  • Best for: audiobooks, dubbing, narrative podcasts, character voices for games

Training & licentie

Not disclosed. ElevenLabs licences professional voice talent, uses public-domain audiobooks and crowd-sourced opt-in voice contributions; commercial recordings are excluded per their public statements.

Licentie: Proprietary commercial SaaS. Commercial use of generated audio is permitted on paid plans; voice clones remain customer property.

Veiligheidstests: AI Speech Classifier offered for free; mandatory consent statement and KYC for Professional Voice Clones; inaudible watermark on outputs; mis-use bans after 2024 deepfake incidents.

Bekende beperkingen

  • Higher latency than v2 Turbo or Cartesia Sonic
  • Tag interpretation occasionally inconsistent in alpha
  • Hard refusal for likeness of named public figures without verified consent
  • Closed weights, no on-premise deployment
  • Pricing per character is among the highest in the market
04

Prijzen

Momenteel niet beschikbaar: dit model is gedeactiveerd. Er is momenteel geen prijs voor dit model, dus het kan niet worden uitgevoerd.

05

API

Roep ElevenLabs v3 (alpha) aan met je Railwail API-sleutel. Gebruik deze model-ID in het verzoek:

Geen geverifieerd API-voorbeeld

De openbare API geeft een ander invoerformaat door dan dit model nodig heeft. Gebruik de playground hierboven.

06

Specificaties

Model-ID
eleven-v3
Ontwikkelaar
ElevenLabs
Invoer
Tekst
Uitvoer
Audio
Uitvoerformaten
MP3, WAV, OPUS
Modelgrootte
Undisclosed
Licentie
Proprietary commercial SaaS. Commercial use of generated audio is permitted on paid plans; voice clones remain customer property.
Catalogusitem bijgewerkt
25 juni 2026

Invoerparameters

Invoeren en instellingen uit het invoerschema van het model. Het voorbeeld in de API-sectie toont welke daarvan de API accepteert.

  • inputVerplicht

    Text to convert to speech

    Type: Tekst
    Standaard: –
    Toegestane waarden: tot 4.000 tekens
  • speed
    Type: Getal
    Standaard: 1
    Toegestane waarden: 0,5 tot 2
  • voice
    Type: Keuze
    Standaard: Rachel
    Toegestane waarden: Rachel, Adam, Antoni, Bella, Domi, Elli, Josh of Sam
  • response_format
    Type: Keuze
    Standaard: mp3
    Toegestane waarden: mp3, wav of opus

Tags

  • elevenlabs
  • tts
  • expressive
  • alpha
  • per-character
07

Gebruiksscenario's

Waarvoor het wordt gebruikt

  • Audiobook narration with emotion
  • Multilingual film and series dubbing
  • Character voices for video games
  • Narrative podcasts and radio drama
  • Accessibility tools and screen readers
08

Veelgestelde vragen

Wat is ElevenLabs v3 (alpha)?

ElevenLabs v3 (alpha) is een model van ElevenLabs in de categorie Tekst-naar-spraak. Het staat in de Railwail-catalogus, maar kan momenteel niet worden uitgevoerd.

Hoeveel kost ElevenLabs v3 (alpha) op Railwail?

ElevenLabs v3 (alpha) kan momenteel niet op Railwail worden uitgevoerd, dus er is geen huidige prijs. Beschikbare alternatieven met prijzen staan verderop op deze pagina.

Welke instellingen ondersteunt ElevenLabs v3 (alpha)?

Volgens het invoerschema kent ElevenLabs v3 (alpha) deze parameters: input (tot 4.000 tekens), speed (0,5 tot 2), voice (Rachel, Adam, Antoni, Bella, Domi, Elli, Josh of Sam) en response_format (mp3, wav of opus).

Hoe snel is ElevenLabs v3 (alpha)?

Er zijn nog niet genoeg gemeten runs van ElevenLabs v3 (alpha) op Railwail om een uitvoeringstijd op te geven. Dit hangt af van de invoer, de instellingen en de belasting bij de provider.

Is ElevenLabs v3 (alpha) beter dan AudioLDM 2?

Dat hangt van de taak af. ElevenLabs v3 (alpha) (ElevenLabs) en AudioLDM 2 (AudioLDM) zijn beide modellen in de categorie Tekst-naar-spraak. De vergelijkingspagina toont hun prijzen en specificaties naast elkaar.

ElevenLabs v3 (alpha) en AudioLDM 2 vergelijken

Kan ik ElevenLabs v3 (alpha) nu gebruiken?

Momenteel niet beschikbaar: dit model is gedeactiveerd. De pagina blijft online; beschikbare alternatieven uit dezelfde categorie staan verderop.

Alle modellen via één API

Één API-sleutel voor elk model op Railwail. Gebruik wordt afgerekend via vooraf gekochte credits, 1 credit = US$ 0,01.