LLM API Cost Calculator

Compare the monthly cost of 143+ LLM APIs from a single screen. Drag the sliders, see your exact bill.

Last updated

Pricing AI workloads is harder than it should be. Every provider quotes a different unit — per 1K tokens, per 1M tokens, per call, per second of audio — and most published rate cards leave out fixed costs, batching discounts, and image surcharges. This calculator normalises 409+ models from Railwail's catalogue so you can compare them apples-to-apples in seconds.

How the math works

Each model has a per-1K-token input price and a per-1K-token output price, stored in credits (1 credit = $0.01). Your monthly bill is simply input_tokens × input_price + output_tokens × output_price. We surface the per-1M numbers in the table because most providers quote that unit publicly, but the underlying math runs on per-1K. Fixed per-call costs (image generation, video, audio) don't apply to token-based models and are excluded from this view — see /pricing for per-call workloads.

Switch providers, keep the same code

All models in the table speak the same OpenAI-compatible API on Railwail. Once you've identified a cheaper option, swapping is one environment-variable change — no SDK rewrite, no retry logic. Compare any two models head-to-head on the /compare page, or browse the full catalogue at /models.

Switching from GPT-5.5 to OpenAI text-embedding-3-small would save you ~100% on a 1M-input + 1M-output workload — same monthly traffic, same prompts, just a different model behind the API.

100K
tokens / month

The prompt + context you send to the model each month.

100K
tokens / month

The total response volume the model generates.

Token-based categories only — image, video, audio use per-call pricing.

Filter the table to one upstream provider.

Showing 143 models.Cheapest match: BLIP at $0.00030/month.
Action
——n/a512Try
——n/a8KTry
Bio_ClinicalBERT
huggingface
——n/a—Try
——n/a—Try
——n/a200KTry
——n/a200KTry
——n/a200KTry
——n/a256KTry
——n/a131KTry
——n/a1KTry
——n/a2.1MTry
——n/a1.0MTry
——n/a1.0MTry
——n/a1.0MTry
——n/a131KTry
——n/a—Try
——n/a4.1MTry
——n/a8KTry
——n/a200KTry
——n/a512Try
——n/a512Try
——n/a32KTry
——n/a256KTry
——n/a256KTry
——n/a—Try
——n/a—Try
——n/a512Try
——n/a512Try
——n/a200KTry
——n/a—Try
——n/a—Try
——n/a2KTry
——n/a8KTry
——n/a4KTry
——n/a128KTry
——n/a128KTry
——n/a512Try
——n/a16KTry
——n/a128KTry
——n/a64KTry
——n/a64KTry
——n/a1.1MTry
——n/a1.1MTry
——n/a1.1MTry
——n/a8KTry
——n/a1.0MTry
——n/a1.0MTry
——n/a1.0MTry
——n/a256KTry
——n/a8KTry
——n/a131KTry
——n/a—Try
——n/a128KTry
——n/a512Try
——n/a127KTry
——n/a127KTry
——n/a131KTry
——n/a33KTry
——n/a131KTry
——n/a131KTry
——n/a33KTry
——n/a128KTry
——n/a16KTry
——n/a128KTry
——n/a4KTry
——n/a512Try
——n/a4KTry
——n/a16KTry
——n/a32KTry
——n/a33KTry
——n/a512Try
——$0.0003—Try
——$0.0003—Try
——$0.0003—Try
——$0.0006—Try
——$0.0012—Try
——$0.0012—Try
——$0.0012—Try
——$0.00128KTry
——$0.0012—Try
——$0.00124KTry
——$0.0014—Try
——$0.0015—Try
——$0.001533KTry
——$0.0020—Try
$0.024—$0.00248KTry
——$0.0024—Try
——$0.003016KTry
——$0.0039—Try
——$0.0050—Try
——$0.006516KTry
——$0.0070131KTry
——$0.007416KTry
——$0.00864KTry
——$0.0094—Try
——$0.0118KTry
——$0.013—Try
$0.156—$0.0168KTry
——$0.018—Try
——$0.020—Try
——$0.022—Try
——$0.02516KTry
——$0.028—Try
$0.060$0.300$0.036128KTry
——$0.04116KTry
——$0.041—Try
——$0.046—Try
——$0.05016KTry
——$0.050—Try
——$0.065—Try
$0.120$0.600$0.0728KTry
$0.180$0.720$0.090128KTry
$0.180$0.720$0.090128KTry
——$0.09216KTry
——$0.094—Try
——$0.1024KTry
$0.240$1.50$0.174400KTry
$0.360$1.44$0.1801.0MTry
$0.360$1.44$0.1801.0MTry
——$0.21616KTry
$0.300$2.40$0.270400KTry
$0.600$3.60$0.4201.0MTry
$0.900$5.40$0.630400KTry
$1.58$4.75$0.6341.0MTry
$1.32$5.28$0.660200KTry
$1.32$5.28$0.660200KTry
$1.20$6.00$0.720200KTry
——$1.1916KTry
$2.40$9.60$1.201.0MTry
$2.40$9.60$1.20200KTry
$1.50$12.00$1.351.0MTry
$1.50$12.00$1.35400KTry
$2.40$12.00$1.441.0MTry
$3.00$12.00$1.50128KTry
$3.00$12.00$1.50128KTry
$2.40$14.40$1.681.0MTry
$3.00$18.00$2.101.1MTry
$3.60$18.00$2.161.0MTry
$4.80$24.00$2.881.0MTry
$6.00$30.00$3.601.0MTry
$6.00$30.00$3.601.0MTry
$6.00$36.00$4.20400KTry
$12.00$60.00$7.201.0MTry

Frequently asked questions

How is the monthly cost calculated?

For each model we apply (input_tokens × input_price + output_tokens × output_price) / 1000. Prices are stored in credits per 1K tokens (1 credit = $0.01); the table shows the USD per-1M equivalent because providers quote that unit publicly.

Why is OpenAI more expensive than open-weights alternatives?

Closed-frontier models bake R&D, alignment work, and hosted inference into a single price. Open-weights models (Llama, Mistral, DeepSeek) push the inference cost to hosting providers competing on margin, which is why per-token prices can be 5-50× lower for comparable quality.

What's the difference between input and output cost?

Input tokens are the prompt you send; output tokens are the model's response. Output is typically 2-5× more expensive because generation is sequential and dominates GPU time. Reasoning models (o1, DeepSeek R1) charge for hidden "thinking" tokens as output, which is the main cost driver.

Does this include batch discounts or volume pricing?

No — list price only. Most providers offer 25-50% discounts for batch endpoints and enterprise contracts; use the table as a ceiling, not a floor.

Sign up to start using these models

One OpenAI-compatible endpoint, every model in the table above. Free credits to start, transparent per-token pricing thereafter.