SeamlessM4T v2 Large (Speech)
seamlessm4t-v2-large-speechMeta SeamlessM4T v2 Large speech mode. Speech-to-speech, speech-to-text, and text-to-speech translation across 100+ languages in a single unified model.
- Price
- ≈ $0.0012/run
- Input → output
- Audio → Text
- Developer
- Community
- Updated
- September 23, 2026
Playground
Try SeamlessM4T v2 Large (Speech)
No input form
No input form for this model yet
Its inputs are not documented yet. So that no run fails on a wrong input, we don't offer a form here. Pick a comparable model instead.
Examples
- Length: 0:02
Output (JSON, shortened)
{ "text_output": "Mon animal préféré est l'éléphant.", "audio_output": null }- Length: 0:02
About SeamlessM4T v2 Large (Speech)
SeamlessM4T v2 Large (Speech) is a model by Community in the Speech-to-text category. On Railwail, SeamlessM4T v2 Large (Speech) costs ≈ $0.0012 per run.
Pricing
| Typical run (≈ 1 s on L40S) | $0.0012 per run |
|---|---|
| GPU time (L40S) | $0.00117 per GPU second |
- Billed by the GPU time the run actually takes. When the run starts, 3× the typical price is reserved from your balance and settled afterwards.
- 1 credit = $0.01
Cost calculator
Price calculator
Typical according to the provider: about 1 s
Total
$0.12
12 credits
Per run
$0.0012 · 0.12 credits
Billed by the actual GPU time; this is an estimate.
API
No verified API example
The inputs of this model are not documented yet.
Specifications
- Model ID
seamlessm4t-v2-large-speech- Developer
- Community
- Category
- Speech-to-text
- Input
- Audio
- Output
- Text
- Billing
- By usage (tokens or GPU time)
- Catalog entry updated
- September 23, 2026
Tags
- replicate
- translation
- meta
- open-weights
- speech-to-speech
- transcription
- seamless
Use cases
Frequently asked questions
What is SeamlessM4T v2 Large (Speech)?
SeamlessM4T v2 Large (Speech) is a model by Community in the Speech-to-text category.
How much does SeamlessM4T v2 Large (Speech) cost on Railwail?
On Railwail, SeamlessM4T v2 Large (Speech) costs ≈ $0.0012 per run. You are charged for what each request actually uses. Usage is paid from prepaid credits; 1 credit equals $0.01.
How fast is SeamlessM4T v2 Large (Speech)?
There are not enough measured runs of SeamlessM4T v2 Large (Speech) on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.
Is SeamlessM4T v2 Large (Speech) better than Incredibly Fast Whisper?
That depends on the task. SeamlessM4T v2 Large (Speech) (Community) and Incredibly Fast Whisper (Community) are both models in the Speech-to-text category. The comparison page shows their prices and specifications side by side.
Compare SeamlessM4T v2 Large (Speech) and Incredibly Fast WhisperComparable models
All in this category- Incredibly Fast WhisperCommunity
Whisper Large v3 wrapped with Hugging Face Transformers optimizations (batched inference, flash attention) for very high throughput. Transcribes hours of audio in minutes on a single GPU. Maintained by Vaibhav Srivastav. Good when you need bulk transcription fast.
≈ $0.0056/run
367 % more expensive per unit
Compare SeamlessM4T v2 Large (Speech) vs. Incredibly Fast Whisper - WhisperOpenAI
OpenAI's Whisper running on Replicate. General-purpose speech recognition trained on 680k hours of multilingual audio. Transcribes and translates 99 languages, robust to accents and background noise, and outputs plain text, segments, or word-level timestamps.
- SeamlessM4TCommunity
Meta's SeamlessM4T multimodal translation model. Takes speech or text input and produces transcription or translation across about 100 languages, including speech-to-text and speech-to-speech. One model covers ASR plus cross-lingual translation without chaining separate systems.
All models through one API
One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.