Gemini 3 Flash

MultimodaalAfgelopenBeschikbaar
van Google DeepMindModel-ID: gemini-3-flash

Google's April 2026 fast multimodal model. Combines Gemini 3 Pro's reasoning with Flash-tier latency and price. Default model in the Gemini app.

Prijs · 1M in/uit
US$ 0,60 / US$ 3,60
Context
1.048.576 tokens
Max. uitvoer
65.536 tokens
Invoer → Uitvoer
Tekst + Afbeelding + Audio + Video → Tekst
Ontwikkelaar
Google DeepMind
Bijgewerkt
23 september 2026

De provider faset dit model uit.

01

Playground

Gemini 3 Flash proberen

Chat

US$ 0,60/1M in
Gemini 3 Flash proberen

Stuur een bericht. Het antwoord komt volledig binnen zodra het model klaar is (geen streaming).

Max. antwoordlengte (tokens)

Deze uitvoering

maximaal US$ 0,0037 · 0,37 credits gereserveerd

Gefactureerd worden de werkelijk gebruikte tokens; het ongebruikte deel van de reservering wordt terugbetaald.

Nieuw hier?

10 gratis credits (US$ 0,10) wanneer je je aanmeldt met Google

Bruikbaar 24 uur na aanmelding, tot 5 uitvoeringen per dag en maximaal 2 credits per uitvoering. Andere aanmeldmethoden starten zonder credits. Voldoende voor 27 uitvoeringen van dit model.

02

Over Gemini 3 Flash

SamengevatPer 23 september 2026

Gemini 3 Flash is een model van Google DeepMind in de categorie Multimodaal. Op Railwail kost Gemini 3 Flash US$ 0,60 per 1M invoertokens en US$ 3,60 per 1M uitvoertokens. Het contextvenster bevat 1.048.576 tokens, en een antwoord kan tot 65.536 tokens lang zijn.

Announced April 22, 2026, Gemini 3 Flash brings Pro-grade reasoning to the Flash latency tier. 1M-token context, fully multimodal (text, image, audio, video), 65K max output. The default model in the Gemini app and AI Mode in Search. PhD-level reasoning on common benchmarks at a fraction of the cost of 3.1 Pro. Recommended for high-throughput agentic workflows, real-time multimodal chat, RAG and consumer applications.

Achtergrond

Over Google DeepMind

Opgericht 2010 · Mountain View, USA / London, UK

Google DeepMind is the merged AI research organisation formed in April 2023 by combining Google Brain with DeepMind. Demis Hassabis leads the unit as CEO. Flash variants have been Google's high-throughput tier since Gemini 1.5 Flash (May 2024), with Gemini 2.0 Flash (December 2024), 2.5 Flash (mid-2025) and Gemini 3 Flash (April 2026) representing the progression. DeepMind's seminal papers include 'Attention Is All You Need' (2017), AlphaGo (2016), AlphaFold (2018-2021, Nobel Prize 2024) and the Gemini Technical Report.

Google DeepMind bezoeken

Architectuur

Sparse Mixture-of-Experts Transformer (multimodal, latency-optimized)

Gemini 3 Flash was announced April 22, 2026 as the default Flash-tier model and the new default model in the Gemini app and AI Mode in Search. It is a natively multimodal Sparse MoE Transformer engineered to combine Gemini 3 Pro's reasoning quality with Flash-grade latency, efficiency and cost. Pretraining used Google's TPU v6e infrastructure on a multi-trillion-token mixture of web text, code, books, image-text pairs, audio and video frames. Post-training combined supervised fine-tuning, RLHF, RL against verifiable rewards and distillation from larger Gemini 3.1 Pro teacher models. The architecture preserves Gemini's native multimodality across text, image, audio and video, the full tool-use API and Search grounding, while running at a fraction of Pro pricing. Gemini 3 Flash is the recommended default for high-throughput agentic workflows and consumer-facing multimodal chat.

Parameters
Undisclosed (sparse MoE, smaller and sparser than Gemini 3.1 Pro)
Context
1.048.576 tokens

Mogelijkheden

  • Pro-grade reasoning at Flash latency
  • 1,048,576 token context window
  • Natively multimodal: text, image, audio and video
  • Search grounding and Code Execution built into the API
  • Function calling, JSON schema and parallel tool calls
  • Default model in the Gemini app and AI Mode in Search
  • PhD-level reasoning on common benchmarks
  • Available via Vertex AI, AI Studio, Gemini Enterprise, Antigravity and the Gemini app
  • Strong long-video understanding (hour-long clips)
  • Cross-lingual fluency across 100+ languages
  • Best for: high-throughput agentic workflows, real-time multimodal chat, RAG, consumer applications.

Training & licentie

Pretrained on a multi-trillion-token mixture of web text, code, books, scientific papers, image-text pairs, audio and video frames. Heavily distilled from larger Gemini 3.1 Pro teacher models. Post-training uses supervised fine-tuning, RLHF and RL against verifiable rewards. Knowledge cutoff in late 2025.

Licentie: Proprietary commercial license via Google AI Studio, Vertex AI and the Gemini app. Free tier available in the Gemini app and AI Mode in Search.

Veiligheidstests: Evaluated under Google DeepMind's Frontier Safety Framework v2 with internal red teams and external evaluators.

Bekende beperkingen

  • Below Gemini 3.1 Pro on the hardest reasoning and long-context benchmarks
  • Smaller context window than 3.1 Pro (1M vs 2M)
  • Vision can misread dense tables and handwriting
  • Region availability is rolling out in 2026
  • Audio output not yet supported
03

Prijzen

Prijzen in US-dollars. Het gebruik wordt in rekening gebracht via vooraf gekochte credits.
InvoerUS$ 0,60 / 1M tokens
UitvoerUS$ 3,60 / 1M tokens
  • Gefactureerd worden de tokens die elke aanvraag daadwerkelijk gebruikt.
  • 1 credit = US$ 0,01

Kostencalculator

Prijscalculator

/ aanvr.
/ aanvr.

Totaal

US$ 0,24

24 credits

Per aanvraag

US$ 0,0024 · 0,24 credits

Elke aanvraag wordt afgerond naar 0,01 credits.

04

API

Roep Gemini 3 Flash aan met je Railwail API-sleutel. Gebruik deze model-ID in het verzoek:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-3-flash",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Stel uw sleutel in als RAILWAIL_API_KEYAPI-sleutel maken
05

Specificaties

Model-ID
gemini-3-flash
Ontwikkelaar
Google DeepMind
Categorie
Multimodaal
Invoer
Tekst, Afbeelding, Audio, Video
Uitvoer
Tekst
Contextvenster
1.048.576 tokens
Max. uitvoer
65.536 tokens
Facturering
Op basis van gebruik (tokens of GPU-tijd)
Levenscyclus
Afgelopen
Modelgrootte
Undisclosed (sparse MoE, smaller and sparser than Gemini 3.1 Pro)
Licentie
Proprietary commercial license via Google AI Studio, Vertex AI and the Gemini app. Free tier available in the Gemini app and AI Mode in Search.
Catalogusitem bijgewerkt
23 september 2026

Invoerparameters

Invoeren en instellingen uit het invoerschema van het model. Het voorbeeld in de API-sectie toont welke daarvan de API accepteert.

  • promptVerplicht

    User message

    Type: Tekst
    Standaard: –
    Toegestane waarden: tot 32.000 tekens
  • top_p
    Type: Getal
    Standaard: 0.95
    Toegestane waarden: 0 tot 1
  • stream
    Type: Ja/Nee
    Standaard: false
    Toegestane waarden: –
  • image_url

    Optional image URL to analyze

    Type: Tekst
    Standaard: –
    Toegestane waarden: –
  • max_tokens
    Type: Geheel getal
    Standaard: 4096
    Toegestane waarden: 1 tot 32.000
  • temperature
    Type: Getal
    Standaard: 1
    Toegestane waarden: 0 tot 2
  • system_prompt

    Optional system instruction

    Type: Tekst
    Standaard: –
    Toegestane waarden: tot 8.000 tekens

Tags

  • google
  • deepmind
  • balanced
  • multimodal
  • low-latency
  • long-context
  • 1m-context
06

Gebruiksscenario's

Waarvoor het wordt gebruikt

  • Default consumer multimodal chat
  • High-throughput agentic workflows
  • Real-time RAG pipelines
  • Long-video summarisation and search
  • Production coding assistants
  • Voice and audio reasoning backends
  • AI Mode in Search and Antigravity workflows
07

Veelgestelde vragen

Wat is Gemini 3 Flash?

Gemini 3 Flash is een model van Google DeepMind in de categorie Multimodaal. Op Railwail kunt u het aanroepen met een API-sleutel via de Railwail-API.

Hoeveel kost Gemini 3 Flash op Railwail?

Op Railwail kost Gemini 3 Flash US$ 0,60 per 1M invoertokens en US$ 3,60 per 1M uitvoertokens. U betaalt voor wat elke aanvraag daadwerkelijk verbruikt. Gebruik wordt betaald met vooraf gekochte credits; 1 credit is gelijk aan US$ 0,01.

Wat is het contextvenster van Gemini 3 Flash?

Het contextvenster van Gemini 3 Flash bevat 1.048.576 tokens. Een antwoord kan tot 65.536 tokens lang zijn.

Hoe snel is Gemini 3 Flash?

Er zijn nog niet genoeg gemeten runs van Gemini 3 Flash op Railwail om een uitvoeringstijd op te geven. Dit hangt af van de invoer, de instellingen en de belasting bij de provider.

Is Gemini 3 Flash beter dan BLIP?

Dat hangt van de taak af. Gemini 3 Flash (Google DeepMind) en BLIP (Salesforce) zijn beide modellen in de categorie Multimodaal. De vergelijkingspagina toont hun prijzen en specificaties naast elkaar.

Gemini 3 Flash en BLIP vergelijken

Kan Gemini 3 Flash afbeeldingen verwerken?

Ja. Gemini 3 Flash accepteert afbeeldingen als invoer naast tekst.

Hoe gebruik ik Gemini 3 Flash via de API?

Maak een Railwail-API-sleutel aan en stuur uw aanvraag met de model-ID gemini-3-flash. Codevoorbeelden voor curl, Python en JavaScript staan in het API-gedeelte van deze pagina.

08

Vergelijkbare modellen

Alle in deze categorie

Gemini 3 Flash via de API gebruiken

Één API-sleutel voor elk model op Railwail. Gebruik wordt afgerekend via vooraf gekochte credits, 1 credit = US$ 0,01.