Riffusion
riffusionStable-Diffusion-based real-time music generator. Operates on spectrogram images then resynthesizes audio, enables seamless transitions and looping.
- Price
- โ $0.0576/run
- Input โ output
- Text โ Audio
- Developer
- Riffusion
- Updated
- September 23, 2026
Playground
Try Riffusion
No input form
No input form for this model yet
Its inputs are not documented yet. So that no run fails on a wrong input, we don't offer a form here. Pick a comparable model instead.
Examples
- Length: 0:05
Prompt
funky synth solo
About Riffusion
Riffusion is a model by Riffusion in the Text-to-speech category. On Railwail, Riffusion costs โ $0.0576 per run.
Pricing
| Typical run (โ 213 s on T4) | $0.0576 per run |
|---|---|
| GPU time (T4) | $0.00027 per GPU second |
- Billed by the GPU time the run actually takes. When the run starts, 3ร the typical price is reserved from your balance and settled afterwards.
- 1 credit = $0.01
Cost calculator
Price calculator
Typical according to the provider: about 213.3 s
Total
$5.76
576 credits
Per run
$0.0576 ยท 5.76 credits
Billed by the actual GPU time; this is an estimate.
API
No verified API example
The inputs of this model are not documented yet.
Specifications
- Model ID
riffusion- Developer
- Riffusion
- Category
- Text-to-speech
- Input
- Text
- Output
- Audio
- Billing
- By usage (tokens or GPU time)
- Catalog entry updated
- September 23, 2026
Tags
- riffusion
- music-generation
- open-weights
- spectrogram
Use cases
Frequently asked questions
What is Riffusion?
Riffusion is a model by Riffusion in the Text-to-speech category.
How much does Riffusion cost on Railwail?
On Railwail, Riffusion costs โ $0.0576 per run. You are charged for what each request actually uses. Usage is paid from prepaid credits; 1 credit equals $0.01.
How fast is Riffusion?
There are not enough measured runs of Riffusion on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.
Is Riffusion better than AudioLDM 2?
That depends on the task. Riffusion (Riffusion) and AudioLDM 2 (Haohe Liu) are both models in the Text-to-speech category. The comparison page shows their prices and specifications side by side.
Compare Riffusion and AudioLDM 2Comparable models
All in this category- AudioLDM 2Haohe Liu
Latent-diffusion model for general-purpose text-to-audio. Generates speech, music, and sound effects with a unified prior.
- Kokoro TTS 82MCommunity
Open-weights 82M-parameter TTS. Punches above its size class on naturalness benchmarks at a fraction of the inference cost of larger models.
- OpenVoice v2Community
MyShell OpenVoice v2. Multilingual zero-shot voice cloning with accurate tone-color reproduction and style/emotion control.
All models through one API
One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.