GPT-5.4 Mini

MultimodaleNuovoDisponibile
di OpenAIID modello: gpt-5-4-mini

OpenAI's efficient mid-tier model. 2x faster than its predecessor, 400k context, approaches GPT-5.4 quality on SWE-Bench Pro at a fraction of the cost.

Prezzo · 1M in/out
0,90 USD / 5,40 USD
Contesto
400.000 token
Max. output
128.000 token
Input → output
Testo + Immagine → Testo
Sviluppatore
OpenAI
Aggiornato
23 settembre 2026
01

Playground

Prova GPT-5.4 Mini

Chat

0,90 USD/1M in
Prova GPT-5.4 Mini

Invia un messaggio. La risposta arriva completa quando il modello ha finito (senza streaming).

Prompt di sistema
Lunghezza massima della risposta (token)

Questa esecuzione

al massimo 0,0056 USD · 0,56 crediti riservati

Fatturato in base ai token effettivamente utilizzati; la parte inutilizzata della prenotazione viene rimborsata.

Nuovo qui?

10 crediti gratuiti (0,10 USD) quando ti iscrivi con Google

Utilizzabile 24 ore dopo l'iscrizione, fino a 5 esecuzioni al giorno e al massimo 2 crediti per esecuzione. Altri metodi di accesso iniziano senza crediti. Sufficiente per 17 esecuzioni di questo modello.

02

Informazioni su GPT-5.4 Mini

RiassuntoA partire da 23 settembre 2026

GPT-5.4 Mini è un modello di OpenAI nella categoria Multimodale. Su Railwail, GPT-5.4 Mini costa 0,90 USD per 1M token di input e 5,40 USD per 1M token di output. La finestra di contesto contiene 400.000 token e una risposta può essere lunga fino a 128.000 token.

Released March 17, 2026, GPT-5.4 mini brings the strengths of GPT-5.4 to a smaller, faster model designed for the subagent era. 400K context, vision input, integrated tool use, and 2x faster latency than GPT-5 mini. Significant gains on coding, reasoning, multimodal understanding and tool use; approaches the full GPT-5.4 on SWE-Bench Pro and OSWorld-Verified. Recommended for subagent workflows, customer-facing chat, coding assistants and high-volume API workloads.

Sfondo

Informazioni su OpenAI

Fondato 2015 · San Francisco, USA

OpenAI was founded in December 2015 as a non-profit AI research organisation and transitioned to a capped-profit structure in 2019. The GPT lineage spans GPT-1 (2018) through GPT-5 (mid-2025) and the GPT-5.x family (2025-2026) which unified the o-series reasoning models with the general-purpose GPT line. GPT-5.4 mini and nano were announced on March 17, 2026 as the small-model tier of the GPT-5.4 generation. OpenAI is backed by Microsoft, Khosla, Andreessen Horowitz, Thrive Capital and Sequoia, with total funding above $60 billion and a 2026 valuation above $300 billion.

Visita OpenAI

Architettura

Unified Transformer (mid-tier, with integrated 'Thinking' tier)

GPT-5.4 mini was announced March 17, 2026 alongside GPT-5.4 nano as the small-model tier of the GPT-5.4 generation. It is a smaller variant of the unified GPT-5.4 architecture, retaining native text + image input, an integrated 'Thinking' tier for reasoning on demand, and the full tool-use API, while running more than 2x faster than GPT-5 mini at significantly lower cost. Pretraining used a similar multi-trillion-token mixture as GPT-5.4 with heavier distillation pressure from larger teacher models. Post-training included supervised fine-tuning, RLHF and reinforcement learning against verifiable rewards on coding, reasoning and tool-use trajectories. On evaluations such as SWE-Bench Pro and OSWorld-Verified, GPT-5.4 mini approaches the performance of full GPT-5.4 while costing roughly one-third as much.

Parametri
Undisclosed (estimated tens of billions of parameters, likely sparse MoE)
Contesto
400.000 token

Capacità

  • 2x faster than GPT-5 mini at lower cost
  • Approaches full GPT-5.4 on SWE-Bench Pro and OSWorld-Verified
  • 400K token context window
  • Native multimodal input: text and images
  • Integrated 'Thinking' tier activates for harder reasoning
  • Native tool use, function calling and parallel tool calls
  • Designed for the subagent era: works well under an orchestrator
  • Strong coding, classification and extraction performance
  • Available in ChatGPT, Codex CLI and the OpenAI API
  • Regional processing endpoints available with 10% uplift
  • Best for: subagent workflows, customer-facing chat, coding assistants, high-volume API workloads.

Addestramento e licenza

Pretrained on a multi-trillion-token mixture of web text, code, scientific papers and licensed data; heavy distillation from larger GPT-5.4 teacher models. Post-training uses supervised fine-tuning, RLHF and RL against verifiable rewards. Knowledge cutoff approximately late 2025.

Licenza: Proprietary commercial license via OpenAI API and Azure OpenAI.

Test di sicurezza: Evaluated under OpenAI's Preparedness Framework with internal and external red-teaming and capability evaluations.

Limitazioni note

  • Below full GPT-5.4 on the hardest agentic and reasoning benchmarks
  • Smaller context window than GPT-5.4 (400K vs 1.05M)
  • No native audio or video input
  • Knowledge cutoff in late 2025
  • Thinking mode adds latency and token cost when enabled
03

Prezzi

Prezzi in dollari USA. L'utilizzo viene addebitato dai crediti prepagati.
Input0,90 USD / 1M token
Output5,40 USD / 1M token
  • Fatturato in base ai token che ogni richiesta utilizza effettivamente.
  • 1 credito = 0,01 USD

Calcolatore di costi

Calcolatore prezzi

/ richiesta
/ richiesta

Totale

0,36 USD

36 crediti

Per richiesta

0,0036 USD · 0,36 crediti

Ogni richiesta viene arrotondata a 0,01 crediti.

04

API

Chiama GPT-5.4 Mini con la tua chiave API Railwail. Usa questo ID modello nella richiesta:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gpt-5-4-mini",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Imposta la tua chiave come RAILWAIL_API_KEYCrea chiave API
05

Specifiche

ID modello
gpt-5-4-mini
Sviluppatore
OpenAI
Categoria
Multimodale
Input
Testo, Immagine
Output
Testo
Finestra di contesto
400.000 token
Output massimo
128.000 token
Fatturazione
In base all'utilizzo (token o tempo GPU)
Dimensione del modello
Undisclosed (estimated tens of billions of parameters, likely sparse MoE)
Licenza
Proprietary commercial license via OpenAI API and Azure OpenAI.
Voce di catalogo aggiornata
23 settembre 2026

Parametri di input

Input e impostazioni dallo schema di input del modello. L'esempio nella sezione API mostra quali di essi l'API accetta.

  • promptObbligatorio

    User message

    Tipo: Testo
    Predefinito: –
    Valori consentiti: fino a 32.000 caratteri
  • top_p
    Tipo: Numero
    Predefinito: 1
    Valori consentiti: 0 a 1
  • stream
    Tipo: Sì/No
    Predefinito: false
    Valori consentiti: –
  • image_url

    Optional image URL to analyze

    Tipo: Testo
    Predefinito: –
    Valori consentiti: –
  • max_tokens
    Tipo: Numero intero
    Predefinito: 4096
    Valori consentiti: 1 a 32.000
  • temperature
    Tipo: Numero
    Predefinito: 1
    Valori consentiti: 0 a 2
  • system_prompt

    Optional system instruction

    Tipo: Testo
    Predefinito: –
    Valori consentiti: fino a 8000 caratteri

Etichette

  • openai
  • balanced
  • cost-efficient
  • vision
  • subagents
  • tools
06

Casi d'uso

A cosa serve

  • Subagent workers under an orchestrator model
  • Production coding assistants and IDE plugins
  • Customer-facing chatbots and triage
  • Large-scale data extraction and classification
  • RAG pipelines with image input
  • Cost-sensitive enterprise APIs
07

Domande frequenti

Cos'è GPT-5.4 Mini?

GPT-5.4 Mini è un modello di OpenAI nella categoria Multimodale. Su Railwail puoi richiamarlo con una chiave API tramite l'API Railwail.

Quanto costa GPT-5.4 Mini su Railwail?

Su Railwail, GPT-5.4 Mini costa 0,90 USD per 1M token di input e 5,40 USD per 1M token di output. Ti viene addebitato ciò che ogni richiesta utilizza effettivamente. L'utilizzo viene pagato con crediti prepagati; 1 credito equivale a 0,01 USD.

Qual è la finestra di contesto di GPT-5.4 Mini?

La finestra di contesto di GPT-5.4 Mini contiene 400.000 token. Una risposta può essere lunga fino a 128.000 token.

Quanto è veloce GPT-5.4 Mini?

Non ci sono ancora abbastanza esecuzioni misurate di GPT-5.4 Mini su Railwail per indicare un tempo di esecuzione. Dipende dall'input, dalle impostazioni e dal carico presso il provider.

GPT-5.4 Mini è migliore di BLIP?

Dipende dall'attività. GPT-5.4 Mini (OpenAI) e BLIP (Salesforce) sono entrambi modelli nella categoria Multimodale. La pagina di confronto mostra i loro prezzi e le specifiche affiancati.

Confronta GPT-5.4 Mini e BLIP

GPT-5.4 Mini può elaborare immagini?

Sì. GPT-5.4 Mini accetta immagini come input oltre al testo.

Come uso GPT-5.4 Mini tramite l'API?

Crea una chiave API Railwail e invia la tua richiesta con l'ID modello gpt-5-4-mini. Gli esempi di codice per curl, Python e JavaScript sono nella sezione API di questa pagina.

08

Modelli comparabili

Tutti in questa categoria

Usa GPT-5.4 Mini tramite l'API

Una chiave API per tutti i modelli su Railwail. L'utilizzo viene addebitato da crediti prepagati, 1 credito = 0,01 USD.