Image models turn a text prompt β and optionally a reference image or a mask β into a finished raster. The category covers everything from photoreal product shots to vector-style illustrations to controllable inpainting and outpainting. You reach for an image model when you need on-brand visuals at scale, when a designer's queue is the bottleneck, or when you want to ship a generative feature inside your own product.
Unlike text models, image generation is billed per-call rather than per-token. On Railwail a single image currently costs from under half a cent (FLUX Schnell, $0.0036) to about eleven cents (Ideogram v3 Quality, $0.108); FLUX 1.1 Pro and Imagen 4 cost $0.048 per image. Higher resolutions and longer step counts cost proportionally more. Some providers expose a separate edit endpoint at a different rate; check the model card before integrating.
The core trade-off is photorealism versus controllability. Diffusion flagships (FLUX 1.1 Pro, Imagen 4, Recraft V3) produce magazine-quality output but ignore detailed compositional instructions about half the time. Smaller models (SDXL, Playground V3, Stable Diffusion 3.5) cost ten times less, render in under two seconds, and let you steer the result with ControlNet, IP-Adapter, or LoRA. For batch production work, the smaller and steerable pipeline almost always wins; for one-off hero shots, reach for the flagship.
Watch out for context dilution in image prompts: most diffusion models cap useful prompt length around 75 tokens, so jamming in twelve adjectives and three style references typically averages them all out instead of stacking them. Write the subject, the action, and the lighting first; everything after the third clause has diminishing influence on the result.
Licensing matters: most providers grant a perpetual commercial-use license on generated images, but a few (FLUX Schnell free tier, some open checkpoints) restrict to non-commercial. The model card spells it out β read it before you put output on a billboard.
Top picks below cover the photorealism flagship, the cheapest workhorse, the longest-prompt model, and the fastest realtime option in the category.