GPT-4o Mini

Texte et chatDisponible
par OpenAIID du modĂšle: gpt-4o-mini

Small, fast, and affordable model for lightweight tasks. Great balance of speed and capability.

Prix · 1M entrée/sortie
0,18 $US / 0,72 $US
Contexte
128.000 tokens
Max. sortie
16.384 tokens
EntrĂ©e → Sortie
Texte → Texte
Temps d'exécution (médiane)
3,6 s
Développeur
OpenAI
01

Playground

Essayer GPT-4o Mini

Chat

0,18 $US/1M entrée
Essayer GPT-4o Mini

Envoyez un message. La réponse arrive complÚte une fois que le modÚle a terminé (sans streaming).

Prompt systĂšme
Longueur max. de la réponse (tokens)

Cette exécution

au maximum 0,0008 $US · 0,08 crédits réservés

Facturation selon les tokens réellement utilisés ; la partie inutilisée de la réservation est remboursée.

Nouveau par ici ?

10 crédits gratuits (0,10 $US) lors de l'inscription avec Google

Utilisable 24 heures aprÚs l'inscription, jusqu'à 5 exécutions par jour et au maximum 2 crédits par exécution. Les autres méthodes de connexion commencent sans crédits. Suffisant pour 125 exécutions de ce modÚle.

02

À propos de GPT-4o Mini

RésuméAu 23 septembre 2026

GPT-4o Mini est un modĂšle de OpenAI dans la catĂ©gorie Texte et chat. Sur Railwail, GPT-4o Mini coĂ»te 0,18 $US par 1M tokens d'entrĂ©e et 0,72 $US par 1M tokens de sortie. La fenĂȘtre de contexte contient 128.000 tokens, et une rĂ©ponse peut faire jusqu'Ă  16.384 tokens.

ArriĂšre-plan

À propos de OpenAI

Fondée 2015 · San Francisco, USA

OpenAI was founded in December 2015 as a non-profit AI research lab by Sam Altman, Elon Musk, Greg Brockman, Ilya Sutskever, Wojciech Zaremba and John Schulman. In 2019 it transitioned to a capped-profit structure (OpenAI LP) to take Microsoft investment, now totalling over $13 billion. Foundational publications include the GPT series papers (GPT-1 through GPT-4), the InstructGPT/RLHF paper (2022) and the GPT-4o System Card (2024). The company ships ChatGPT (launched November 2022), the GPT API, the o-series reasoning models, Sora for video, DALL-E for images and Whisper for speech recognition. Sam Altman remains CEO; Greg Brockman is President. OpenAI's 2025 valuation exceeded $300 billion and the company employs more than 1,500 staff in San Francisco, London, Tokyo, Dublin and other offices. The mission statement focuses on building safe and beneficial AGI for all of humanity.

Visiter OpenAI

Architecture

Decoder-only Transformer (small natively multimodal)

GPT-4o mini was announced in July 2024 as a low-cost, fast replacement for GPT-3.5 Turbo. It is a small natively multimodal Transformer derived from the same training stack as GPT-4o, supporting text and vision input with text output. The model was trained on a multi-trillion-token web-scale corpus including code, books, licensed text and image-text pairs, with a knowledge cutoff of October 2023. Post-training combined supervised fine-tuning with RLHF and direct preference optimisation on smaller curated datasets. OpenAI applied an 'instruction hierarchy' training method to better resist jailbreaks and prompt injections by distinguishing system, developer and user instructions during alignment. At launch GPT-4o mini scored 82% on MMLU and outperformed GPT-3.5 Turbo on the Chatbot Arena leaderboard while costing approximately 60% less. It supports the full GPT-4o feature set including function calling, parallel tool calls, JSON mode, Structured Outputs and vision input. The model is the default backbone for ChatGPT Free as of mid-2024 and is widely used as a router and cheap tool-calling layer in agentic systems. Fine-tuning was opened to developers in late 2024 with both supervised fine-tuning and reinforcement fine-tuning options on the API.

ParamĂštres
Undisclosed (estimated ~8B-20B parameters dense)
Contexte
128.000 tokens

Capacités

  • Very low cost per token (around $0.15 input / $0.60 output per 1M tokens at launch)
  • 128K context window with 16K max output
  • Vision input for images and PDFs
  • Function calling and parallel tool calls
  • Structured Outputs with strict JSON schema
  • 82% MMLU at launch (better than GPT-3.5 Turbo)
  • Trained with instruction hierarchy to resist prompt injection
  • Supervised and reinforcement fine-tuning available
  • Fast time-to-first-token for chat workloads
  • Multilingual coverage but optimised for English
  • Best for: cheap chatbots, classification, large-scale data labeling, agent routers, customer support.

Formation & licence

Pretrained on OpenAI's curated multi-trillion-token mixture of web text, code, books and image-text pairs, with a knowledge cutoff of October 2023. Post-training uses supervised fine-tuning, RLHF and the instruction-hierarchy alignment objective.

Licence: Proprietary, accessible via OpenAI API and Azure OpenAI Service. Commercial use permitted under OpenAI Terms.

Tests de sécurité: Evaluated under OpenAI's Preparedness Framework and System Card process; smaller than GPT-4o so deemed low risk for CBRN uplift but tested against persuasion, cyber and prompt-injection threats.

Limitations connues

  • Weaker reasoning than full GPT-4o on hard math and multi-step problems
  • Higher hallucination rate than larger flagship models
  • Knowledge cutoff October 2023
  • No audio input or output
  • Vision quality below full GPT-4o on complex charts
03

Tarification

Prix en dollars américains. L'utilisation est facturée à partir de crédits prépayés.
Entrée0,18 $US / 1M tokens
Sortie0,72 $US / 1M tokens
  • FacturĂ© selon les tokens que chaque requĂȘte utilise rĂ©ellement.
  • 1 crĂ©dit = 0,01 $US

Calculateur de coûts

Calculatrice de prix

/ req.
/ req.

Total

0,06 $US

6 crédits

Par requĂȘte

0,0006 $US · 0,06 crédits

Chaque requĂȘte est arrondie Ă  0,01 crĂ©dits.

04

API

Appelez GPT-4o Mini avec votre clĂ© API Railwail. Utilisez cet ID de modĂšle dans la requĂȘte :
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gpt-4o-mini",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Définissez votre clé comme RAILWAIL_API_KEYCréer une clé API
05

Spécifications

ID du modĂšle
gpt-4o-mini
Développeur
OpenAI
Catégorie
Texte et chat
Entrée
Texte
Sortie
Texte
FenĂȘtre de contexte
128.000 tokens
Sortie max.
16.384 tokens
Facturation
À l'usage (tokens ou temps GPU)
Temps d'exécution (médiane)
3,6 s7 exécutions terminées sur Railwail au cours des 90 derniers jours
Taille du modĂšle
Undisclosed (estimated ~8B-20B parameters dense)
Licence
Proprietary, accessible via OpenAI API and Azure OpenAI Service. Commercial use permitted under OpenAI Terms.
Entrée du catalogue mise à jour
23 septembre 2026

Étiquettes

  • fast
  • affordable
06

Prompts d'exemple

Exemples du catalogue Railwail. Ils n'ont pas été générés en direct sur cette page.
  • Quick Summary

    Summarize the key points of the Agile methodology in 3 bullet points
    Afficher la réponse d'exemple

    1) Agile emphasizes iterative development, delivering working software in short cycles called sprints rather than one big release. 2) It prioritizes collaboration between cross-functional teams and stakeholders, with regular feedback loops. 3) Agile embraces change—requirements can evolve throughout the project based on user feedback and shifting priorities.

  • Email Draft

    Write a professional email declining a meeting invitation due to a scheduling conflict
    Afficher la réponse d'exemple

    Subject: Re: Meeting Invitation – Schedule Conflict Hi [Name], Thank you for the invitation. Unfortunately, I have a prior commitment at that time and won't be able to attend. I'd love to stay in the loop—could you share the meeting notes afterward? Alternatively, I'm available on Thursday afternoon if rescheduling works for the group. Best regards

07

Cas d'usage

À quoi ça sert

  • Cost-sensitive chatbots
  • Large-scale classification and labeling
  • Agent tool routers
  • Email drafting and summarisation
  • Coding completions in IDEs
  • Customer-service automation
08

Questions fréquemment posées

Qu'est-ce que GPT-4o Mini ?

GPT-4o Mini est un modÚle de OpenAI dans la catégorie Texte et chat. Sur Railwail, vous pouvez l'appeler avec une clé API via l'API Railwail.

Combien coûte GPT-4o Mini sur Railwail ?

Sur Railwail, GPT-4o Mini coĂ»te 0,18 $US par 1M tokens d'entrĂ©e et 0,72 $US par 1M tokens de sortie. Vous ĂȘtes facturĂ© pour ce que chaque requĂȘte utilise rĂ©ellement. L'utilisation est payĂ©e Ă  partir de crĂ©dits prĂ©payĂ©s ; 1 crĂ©dit Ă©quivaut Ă  0,01 $US.

Quelle est la fenĂȘtre de contexte de GPT-4o Mini ?

La fenĂȘtre de contexte de GPT-4o Mini contient 128.000 tokens. Une rĂ©ponse peut faire jusqu'Ă  16.384 tokens.

Quelle est la vitesse de GPT-4o Mini ?

Sur Railwail, le temps d'exécution médian de GPT-4o Mini au cours des 90 derniers jours était 3,6 s, basé sur 7 exécutions terminées.

GPT-4o Mini est-il meilleur que Claude Fable 5.1 ?

Cela dépend de la tùche. GPT-4o Mini (OpenAI) et Claude Fable 5.1 (Anthropic) sont tous deux des modÚles de la catégorie Texte et chat. La page de comparaison affiche leurs prix et spécifications cÎte à cÎte.

Comparer GPT-4o Mini et Claude Fable 5.1

Comment utiliser GPT-4o Mini via l'API ?

CrĂ©ez une clĂ© API Railwail et envoyez votre requĂȘte avec l'ID de modĂšle gpt-4o-mini. Des exemples de code pour curl, Python et JavaScript se trouvent dans la section API de cette page.

09

ModĂšles comparables

Tous dans cette catégorie
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    12,00 $US/1M entrée

    6.567 % plus cher par unité

    Comparer GPT-4o Mini et Claude Fable 5.1
  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    6,00 $US/1M entrée

    3.233 % plus cher par unité

    Comparer GPT-4o Mini et Claude Opus 4.8
  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    4,80 $US/1M entrée

    2.567 % plus cher par unité

    Comparer GPT-4o Mini et Claude Opus 5.5

Utiliser GPT-4o Mini via l'API

Une clé API pour tous les modÚles sur Railwail. L'utilisation est facturée à partir de crédits prépayés, 1 crédit = 0,01 $US.