StyleTTS 2
styletts-2Style-based TTS using diffusion and adversarial training. Human-level naturalness in zero-shot voice synthesis from a 3-5s reference clip.
- Prijs
- ≈ US$ 0,00040/uitvoering
- Invoer → Uitvoer
- Tekst → Audio
- Ontwikkelaar
- Community
- Bijgewerkt
- 23 september 2026
Playground
StyleTTS 2 proberen
Geen invoerformulier
Voor dit model is nog geen invoerformulier beschikbaar
De invoeren zijn nog niet gedocumenteerd. Om te voorkomen dat een uitvoering mislukt door onjuiste invoer, bieden we hier geen formulier aan. Kies in plaats daarvan een vergelijkbaar model.
Examples
- Lengte: 0:13
Prompt
StyleTTS 2 is a text-to-speech model that leverages style diffusion and adversarial training with large speech language models to achieve human-level text-to-speech synthesis.
- Lengte: 0:54
Prompt
If the supply of fruit is greater than the family needs, it may be made a source of income by sending the fresh fruit to the market if there is one near enough, or by preserving, canning, and making jelly for sale. To make such an enterprise a success the fruit and work must be first class. There is magic in the word 'Homemade,' when the product appeals to the eye and the palate; but many careless and incompetent people have found to their sorrow that this word has not magic enough to float inferior goods on the market. As a rule large canning and preserving establishments are clean and have the best appliances, and they employ chemists and skilled labor. The home product must be very good to compete with the attractive goods that are sent out from such establishments. Yet for first-class homemade products there is a market in all large cities. All first-class grocers have customers who purchase such goods.
Over StyleTTS 2
StyleTTS 2 is een model van Community in de categorie Tekst-naar-spraak. Op Railwail kost StyleTTS 2 ≈ US$ 0,00040 per uitvoering.
Prijzen
| Typische uitvoering (≈ 1 s op T4) | US$ 0,00040 per uitvoering |
|---|---|
| GPU-tijd (T4) | US$Â 0,00027 per GPU-seconde |
- Afgerekend wordt de GPU-tijd die de run werkelijk gebruikt. Wanneer de run start, wordt het 3-voudige van de typische prijs van uw saldo gereserveerd en later verrekend.
- 1 credit = US$Â 0,01
Kostencalculator
Prijscalculator
Typisch volgens de provider: ongeveer 1,4 s
Totaal
US$Â 0,04
4 credits
Per run
US$ 0,0004 · 0,04 credits
Afgerekend wordt de werkelijke GPU-tijd; dit is een schatting.
API
Geen geverifieerd API-voorbeeld
De invoer van dit model is nog niet gedocumenteerd.
Specificaties
- Model-ID
styletts-2- Ontwikkelaar
- Community
- Categorie
- Tekst-naar-spraak
- Invoer
- Tekst
- Uitvoer
- Audio
- Facturering
- Op basis van gebruik (tokens of GPU-tijd)
- Catalogusitem bijgewerkt
- 23 september 2026
Tags
- styletts
- tts
- voice-cloning
- diffusion
- open-source
Gebruiksscenario's
Veelgestelde vragen
Wat is StyleTTS 2?
StyleTTS 2 is een model van Community in de categorie Tekst-naar-spraak.
Hoeveel kost StyleTTS 2 op Railwail?
Op Railwail kost StyleTTS 2 ≈ US$ 0,00040 per uitvoering. U betaalt voor wat elke aanvraag daadwerkelijk verbruikt. Gebruik wordt betaald met vooraf gekochte credits; 1 credit is gelijk aan US$ 0,01.
Hoe snel is StyleTTS 2?
Er zijn nog niet genoeg gemeten runs van StyleTTS 2 op Railwail om een uitvoeringstijd op te geven. Dit hangt af van de invoer, de instellingen en de belasting bij de provider.
Is StyleTTS 2 beter dan AudioLDM 2?
Dat hangt van de taak af. StyleTTS 2 (Community) en AudioLDM 2 (Haohe Liu) zijn beide modellen in de categorie Tekst-naar-spraak. De vergelijkingspagina toont hun prijzen en specificaties naast elkaar.
StyleTTS 2 en AudioLDM 2 vergelijkenVergelijkbare modellen
Alle in deze categorie- AudioLDM 2Haohe Liu
Latent-diffusion model for general-purpose text-to-audio. Generates speech, music, and sound effects with a unified prior.
- Kokoro TTS 82MCommunity
Open-weights 82M-parameter TTS. Punches above its size class on naturalness benchmarks at a fraction of the inference cost of larger models.
- OpenVoice v2Community
MyShell OpenVoice v2. Multilingual zero-shot voice cloning with accurate tone-color reproduction and style/emotion control.
Alle modellen via één API
Één API-sleutel voor elk model op Railwail. Gebruik wordt afgerekend via vooraf gekochte credits, 1 credit = US$ 0,01.