GPT-5.4 Mini

MultimodalNeuVerfügbar
von OpenAIModell-ID: gpt-5-4-mini

OpenAI's efficient mid-tier model. 2x faster than its predecessor, 400k context, approaches GPT-5.4 quality on SWE-Bench Pro at a fraction of the cost.

Preis · 1 Mio. In/Out
0,90 $ / 5,40 $
Kontext
400.000 Token
Max. Ausgabe
128.000 Token
Eingabe → Ausgabe
Text + Bild → Text
Entwickler
OpenAI
Aktualisiert
23. September 2026
01

Playground

GPT-5.4 Mini ausprobieren

Chat

0,90 $/1 Mio. In
GPT-5.4 Mini ausprobieren

Schick eine Nachricht. Die Antwort kommt vollständig, sobald das Modell fertig ist (ohne Streaming).

System-Prompt
Max. Antwortlänge (Token)

Dieser Lauf

höchstens 0,0056 $ · 0,56 Credits vorgemerkt

Abgerechnet werden die tatsächlich verbrauchten Token, der Rest der Vormerkung wird erstattet.

Neu hier?

10 Gratis-Credits (0,10 $) bei Anmeldung mit Google

Nutzbar 24 Stunden nach der Anmeldung, bis zu 5 Läufe pro Tag und höchstens 2 Credits je Lauf. Andere Anmeldearten starten ohne Guthaben. Reicht für 17 Läufe dieses Modells.

02

Über GPT-5.4 Mini

Kurz gesagtStand: 23. September 2026

GPT-5.4 Mini ist ein Modell von OpenAI aus der Kategorie Multimodal. Über Railwail kostet GPT-5.4 Mini 0,90 $ je 1 Mio. Input-Token und 5,40 $ je 1 Mio. Output-Token. Das Kontextfenster umfasst 400.000 Token, eine Antwort bis zu 128.000 Token.

Released March 17, 2026, GPT-5.4 mini brings the strengths of GPT-5.4 to a smaller, faster model designed for the subagent era. 400K context, vision input, integrated tool use, and 2x faster latency than GPT-5 mini. Significant gains on coding, reasoning, multimodal understanding and tool use; approaches the full GPT-5.4 on SWE-Bench Pro and OSWorld-Verified. Recommended for subagent workflows, customer-facing chat, coding assistants and high-volume API workloads.

Hintergrund

Über OpenAI

Gegründet 2015 · San Francisco, USA

OpenAI was founded in December 2015 as a non-profit AI research organisation and transitioned to a capped-profit structure in 2019. The GPT lineage spans GPT-1 (2018) through GPT-5 (mid-2025) and the GPT-5.x family (2025-2026) which unified the o-series reasoning models with the general-purpose GPT line. GPT-5.4 mini and nano were announced on March 17, 2026 as the small-model tier of the GPT-5.4 generation. OpenAI is backed by Microsoft, Khosla, Andreessen Horowitz, Thrive Capital and Sequoia, with total funding above $60 billion and a 2026 valuation above $300 billion.

OpenAI besuchen

Architektur

Unified Transformer (mid-tier, with integrated 'Thinking' tier)

GPT-5.4 mini was announced March 17, 2026 alongside GPT-5.4 nano as the small-model tier of the GPT-5.4 generation. It is a smaller variant of the unified GPT-5.4 architecture, retaining native text + image input, an integrated 'Thinking' tier for reasoning on demand, and the full tool-use API, while running more than 2x faster than GPT-5 mini at significantly lower cost. Pretraining used a similar multi-trillion-token mixture as GPT-5.4 with heavier distillation pressure from larger teacher models. Post-training included supervised fine-tuning, RLHF and reinforcement learning against verifiable rewards on coding, reasoning and tool-use trajectories. On evaluations such as SWE-Bench Pro and OSWorld-Verified, GPT-5.4 mini approaches the performance of full GPT-5.4 while costing roughly one-third as much.

Parameter
Undisclosed (estimated tens of billions of parameters, likely sparse MoE)
Kontext
400.000 Token

Fähigkeiten

  • 2x faster than GPT-5 mini at lower cost
  • Approaches full GPT-5.4 on SWE-Bench Pro and OSWorld-Verified
  • 400K token context window
  • Native multimodal input: text and images
  • Integrated 'Thinking' tier activates for harder reasoning
  • Native tool use, function calling and parallel tool calls
  • Designed for the subagent era: works well under an orchestrator
  • Strong coding, classification and extraction performance
  • Available in ChatGPT, Codex CLI and the OpenAI API
  • Regional processing endpoints available with 10% uplift
  • Best for: subagent workflows, customer-facing chat, coding assistants, high-volume API workloads.

Training & Lizenz

Pretrained on a multi-trillion-token mixture of web text, code, scientific papers and licensed data; heavy distillation from larger GPT-5.4 teacher models. Post-training uses supervised fine-tuning, RLHF and RL against verifiable rewards. Knowledge cutoff approximately late 2025.

Lizenz: Proprietary commercial license via OpenAI API and Azure OpenAI.

Sicherheitstests: Evaluated under OpenAI's Preparedness Framework with internal and external red-teaming and capability evaluations.

Bekannte Grenzen

  • Below full GPT-5.4 on the hardest agentic and reasoning benchmarks
  • Smaller context window than GPT-5.4 (400K vs 1.05M)
  • No native audio or video input
  • Knowledge cutoff in late 2025
  • Thinking mode adds latency and token cost when enabled
03

Preise

Preise in US-Dollar. Abgerechnet wird über vorab gekaufte Credits.
Eingabe0,90 $ / 1 Mio. Token
Ausgabe5,40 $ / 1 Mio. Token
  • Abgerechnet werden die Token, die jede Anfrage tatsächlich verbraucht.
  • 1 Credit = 0,01 $

Kostenrechner

Preisrechner

/ Anfr.
/ Anfr.

Gesamt

0,36 $

36 Credits

Je Anfrage

0,0036 $ · 0,36 Credits

Jede Anfrage wird auf 0,01 Credits aufgerundet.

04

API

Rufe GPT-5.4 Mini mit deinem Railwail-API-Schlüssel auf. Diese Modell-ID gehört in die Anfrage:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gpt-5-4-mini",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Schlüssel als RAILWAIL_API_KEY setzenAPI-Schlüssel erstellen
05

Spezifikationen

Modell-ID
gpt-5-4-mini
Entwickler
OpenAI
Kategorie
Multimodal
Eingabe
Text, Bild
Ausgabe
Text
Kontextfenster
400.000 Token
Max. Ausgabe
128.000 Token
Abrechnung
Nach Verbrauch (Token bzw. GPU-Zeit)
Modellgröße
Undisclosed (estimated tens of billions of parameters, likely sparse MoE)
Lizenz
Proprietary commercial license via OpenAI API and Azure OpenAI.
Katalogeintrag aktualisiert
23. September 2026

Eingabeparameter

Eingaben und Einstellungen laut Eingabeschema des Modells. Welche davon die API annimmt, zeigt das Beispiel im Abschnitt API.

  • promptPflicht

    User message

    Typ: Text
    Standard: –
    Erlaubte Werte: bis 32.000 Zeichen
  • top_p
    Typ: Zahl
    Standard: 1
    Erlaubte Werte: 0 bis 1
  • stream
    Typ: Ja/Nein
    Standard: false
    Erlaubte Werte: –
  • image_url

    Optional image URL to analyze

    Typ: Text
    Standard: –
    Erlaubte Werte: –
  • max_tokens
    Typ: Ganzzahl
    Standard: 4096
    Erlaubte Werte: 1 bis 32.000
  • temperature
    Typ: Zahl
    Standard: 1
    Erlaubte Werte: 0 bis 2
  • system_prompt

    Optional system instruction

    Typ: Text
    Standard: –
    Erlaubte Werte: bis 8.000 Zeichen

Schlagwörter

  • openai
  • balanced
  • cost-efficient
  • vision
  • subagents
  • tools
06

Einsatzgebiete

Wofür es genutzt wird

  • Subagent workers under an orchestrator model
  • Production coding assistants and IDE plugins
  • Customer-facing chatbots and triage
  • Large-scale data extraction and classification
  • RAG pipelines with image input
  • Cost-sensitive enterprise APIs
07

Häufige Fragen

Was ist GPT-5.4 Mini?

GPT-5.4 Mini ist ein Modell von OpenAI aus der Kategorie Multimodal. Über Railwail lässt es sich mit einem API-Schlüssel über die Railwail-API aufrufen.

Was kostet GPT-5.4 Mini bei Railwail?

Über Railwail kostet GPT-5.4 Mini 0,90 $ je 1 Mio. Input-Token und 5,40 $ je 1 Mio. Output-Token. Abgerechnet wird, was jede Anfrage tatsächlich verbraucht. Bezahlt wird mit vorab gekauften Credits; 1 Credit entspricht 0,01 $.

Wie groß ist das Kontextfenster von GPT-5.4 Mini?

Das Kontextfenster von GPT-5.4 Mini umfasst 400.000 Token. Eine Antwort kann bis zu 128.000 Token lang sein.

Wie schnell ist GPT-5.4 Mini?

Für GPT-5.4 Mini gibt es bei Railwail noch zu wenige gemessene Läufe, um eine Laufzeit anzugeben. Sie hängt von der Eingabe, den Einstellungen und der Auslastung beim Anbieter ab.

Ist GPT-5.4 Mini besser als BLIP?

Das hängt von der Aufgabe ab. GPT-5.4 Mini (OpenAI) und BLIP (Salesforce) sind beide Modelle aus der Kategorie Multimodal. Die Vergleichsseite zeigt Preise und Spezifikationen nebeneinander.

GPT-5.4 Mini und BLIP vergleichen

Kann GPT-5.4 Mini Bilder verarbeiten?

Ja. GPT-5.4 Mini nimmt neben Text auch Bilder als Eingabe an.

Wie nutze ich GPT-5.4 Mini über die API?

Erstelle einen Railwail-API-Schlüssel und sende deine Anfrage mit der Modell-ID gpt-5-4-mini. Codebeispiele für curl, Python und JavaScript stehen im Abschnitt API auf dieser Seite.

08

Vergleichbare Modelle

Alle dieser Kategorie

GPT-5.4 Mini über die API nutzen

Ein API-Schlüssel für alle Modelle auf Railwail. Abgerechnet wird über vorab gekaufte Credits, 1 Credit = 0,01 $.