GPT-4.1
OpenAI's newest flagship model. Improved reasoning, instruction following, and coding over GPT-4o.
OpenAI develops GPT-4o, DALL-E 3, Whisper, and the o1 reasoning family β the most widely-adopted commercial AI models on the market.
Access every OpenAI model through Railwail's OpenAI-compatible API.
24 models
OpenAI's newest flagship model. Improved reasoning, instruction following, and coding over GPT-4o.
OpenAI's most capable multimodal model. Excellent for complex reasoning, coding, and creative tasks.
OpenAI's unified flagship combining GPT and o-series reasoning into one model. 1M context, multimodal, top SWE-Bench Pro and OSWorld scores.
OpenAI's efficient mid-tier model. 2x faster than its predecessor, 400k context, approaches GPT-5.4 quality on SWE-Bench Pro at a fraction of the cost.
OpenAI's current flagship chat model (released April 2026). Strongest general reasoning, coding and tool use in the GPT-5 line, with vision input and a large context window.
OpenAI's reasoning model optimized for STEM tasks, coding, and math. Uses chain-of-thought reasoning.
OpenAI's highest-quality embedding model. Returns 3072-dim vectors by default and supports reducing dimensions via the dimensions parameter. Outperforms text-embedding-3-small and the older ada-002 on MTEB and multilingual MIRACL retrieval benchmarks, for cases where accuracy matters more than cost.
OpenAI's small, low-cost embedding model. Returns 1536-dim vectors by default and supports shortening output dimensions via the dimensions parameter without retraining. Replaced text-embedding-ada-002 with better retrieval quality at a fraction of the price, and is the default choice for general-purpose semantic search and RAG.
OpenAI's Whisper running on Replicate. General-purpose speech recognition trained on 680k hours of multilingual audio. Transcribes and translates 99 languages, robust to accents and background noise, and outputs plain text, segments, or word-level timestamps.
OpenAI's latest image generation model with better instruction following and adherence to prompts
OpenAI's state-of-the-art image generation model. Create and edit images from text with strong instruction following, sharp text rendering, and detailed editing.
OpenAI's fastest model for high-quality, everyday image generation. Generate and edit images from text and image inputs with strong instruction following and sharp text rendering.
OpenAI's most capable image model, built for workflows where editing precision matters most. Generate and edit images from text and image inputs with strong instruction following, sharp text rendering, and detailed control.
Small, fast, and affordable model for lightweight tasks. Great balance of speed and capability.
Smaller, faster, cheaper member of OpenAI's GPT-5 family. Tuned for high-throughput chat, classification and extraction where the full flagship is overkill.
OpenAI GPT-5.1 chat model (November 2025). An earlier GPT-5 point release kept available for compatibility. Good general-purpose reasoning and coding.
OpenAI's smallest and cheapest GPT-5.4 variant. Built for high-volume classification, extraction and coding subagents at edge-grade latency.
OpenAI's most capable model, built for the hardest end-to-end work. Reasoning, text and image input, function calling and tool use, 1,050,000-token context window and up to 128,000 output tokens. Knowledge cutoff April 30, 2026.
OpenAI's most efficient model for focused, high-volume tasks. Reasoning, text and image input, function calling and tool use, 1,050,000-token context window and up to 128,000 output tokens. Knowledge cutoff May 18, 2026.
OpenAI model built to power complex coding and agentic workflows. Reasoning, text and image input, function calling and tool use, 1,050,000-token context window and up to 128,000 output tokens. Knowledge cutoff April 20, 2026.
OpenAI's o3 reasoning model. Spends compute on a private chain of thought before answering, strong at math, science and hard coding problems that benefit from deliberate reasoning.
OpenAI's o4-mini reasoning model. A cost-efficient reasoning model that trades some depth for much lower price and latency, good for high-volume math and code tasks.
OpenAI's text-to-speech model. Six built-in voices with natural intonation.
OpenAI's high-definition TTS model. Better quality for production use cases.
Listed for reference. These models have no verified price or are no longer served, so the API does not run them.
Free credits on sign-up. No credit card required. Access OpenAI and 27+ other providers through a single API.