Gemini 3 Flash

MultimodaleDeprecatoDisponibile
di Google DeepMindID modello: gemini-3-flash

Google's April 2026 fast multimodal model. Combines Gemini 3 Pro's reasoning with Flash-tier latency and price. Default model in the Gemini app.

Prezzo · 1M in/out
0,60 USD / 3,60 USD
Contesto
1.048.576 token
Max. output
65.536 token
Input → output
Testo + Immagine + Audio + Video → Testo
Sviluppatore
Google DeepMind
Aggiornato
23 settembre 2026

Il provider sta eliminando gradualmente questo modello.

01

Playground

Prova Gemini 3 Flash

Chat

0,60 USD/1M in
Prova Gemini 3 Flash

Invia un messaggio. La risposta arriva completa quando il modello ha finito (senza streaming).

Lunghezza massima della risposta (token)

Questa esecuzione

al massimo 0,0037 USD · 0,37 crediti riservati

Fatturato in base ai token effettivamente utilizzati; la parte inutilizzata della prenotazione viene rimborsata.

Nuovo qui?

10 crediti gratuiti (0,10 USD) quando ti iscrivi con Google

Utilizzabile 24 ore dopo l'iscrizione, fino a 5 esecuzioni al giorno e al massimo 2 crediti per esecuzione. Altri metodi di accesso iniziano senza crediti. Sufficiente per 27 esecuzioni di questo modello.

02

Informazioni su Gemini 3 Flash

RiassuntoA partire da 23 settembre 2026

Gemini 3 Flash è un modello di Google DeepMind nella categoria Multimodale. Su Railwail, Gemini 3 Flash costa 0,60 USD per 1M token di input e 3,60 USD per 1M token di output. La finestra di contesto contiene 1.048.576 token e una risposta può essere lunga fino a 65.536 token.

Announced April 22, 2026, Gemini 3 Flash brings Pro-grade reasoning to the Flash latency tier. 1M-token context, fully multimodal (text, image, audio, video), 65K max output. The default model in the Gemini app and AI Mode in Search. PhD-level reasoning on common benchmarks at a fraction of the cost of 3.1 Pro. Recommended for high-throughput agentic workflows, real-time multimodal chat, RAG and consumer applications.

Sfondo

Informazioni su Google DeepMind

Fondato 2010 · Mountain View, USA / London, UK

Google DeepMind is the merged AI research organisation formed in April 2023 by combining Google Brain with DeepMind. Demis Hassabis leads the unit as CEO. Flash variants have been Google's high-throughput tier since Gemini 1.5 Flash (May 2024), with Gemini 2.0 Flash (December 2024), 2.5 Flash (mid-2025) and Gemini 3 Flash (April 2026) representing the progression. DeepMind's seminal papers include 'Attention Is All You Need' (2017), AlphaGo (2016), AlphaFold (2018-2021, Nobel Prize 2024) and the Gemini Technical Report.

Visita Google DeepMind

Architettura

Sparse Mixture-of-Experts Transformer (multimodal, latency-optimized)

Gemini 3 Flash was announced April 22, 2026 as the default Flash-tier model and the new default model in the Gemini app and AI Mode in Search. It is a natively multimodal Sparse MoE Transformer engineered to combine Gemini 3 Pro's reasoning quality with Flash-grade latency, efficiency and cost. Pretraining used Google's TPU v6e infrastructure on a multi-trillion-token mixture of web text, code, books, image-text pairs, audio and video frames. Post-training combined supervised fine-tuning, RLHF, RL against verifiable rewards and distillation from larger Gemini 3.1 Pro teacher models. The architecture preserves Gemini's native multimodality across text, image, audio and video, the full tool-use API and Search grounding, while running at a fraction of Pro pricing. Gemini 3 Flash is the recommended default for high-throughput agentic workflows and consumer-facing multimodal chat.

Parametri
Undisclosed (sparse MoE, smaller and sparser than Gemini 3.1 Pro)
Contesto
1.048.576 token

Capacità

  • Pro-grade reasoning at Flash latency
  • 1,048,576 token context window
  • Natively multimodal: text, image, audio and video
  • Search grounding and Code Execution built into the API
  • Function calling, JSON schema and parallel tool calls
  • Default model in the Gemini app and AI Mode in Search
  • PhD-level reasoning on common benchmarks
  • Available via Vertex AI, AI Studio, Gemini Enterprise, Antigravity and the Gemini app
  • Strong long-video understanding (hour-long clips)
  • Cross-lingual fluency across 100+ languages
  • Best for: high-throughput agentic workflows, real-time multimodal chat, RAG, consumer applications.

Addestramento e licenza

Pretrained on a multi-trillion-token mixture of web text, code, books, scientific papers, image-text pairs, audio and video frames. Heavily distilled from larger Gemini 3.1 Pro teacher models. Post-training uses supervised fine-tuning, RLHF and RL against verifiable rewards. Knowledge cutoff in late 2025.

Licenza: Proprietary commercial license via Google AI Studio, Vertex AI and the Gemini app. Free tier available in the Gemini app and AI Mode in Search.

Test di sicurezza: Evaluated under Google DeepMind's Frontier Safety Framework v2 with internal red teams and external evaluators.

Limitazioni note

  • Below Gemini 3.1 Pro on the hardest reasoning and long-context benchmarks
  • Smaller context window than 3.1 Pro (1M vs 2M)
  • Vision can misread dense tables and handwriting
  • Region availability is rolling out in 2026
  • Audio output not yet supported
03

Prezzi

Prezzi in dollari USA. L'utilizzo viene addebitato dai crediti prepagati.
Input0,60 USD / 1M token
Output3,60 USD / 1M token
  • Fatturato in base ai token che ogni richiesta utilizza effettivamente.
  • 1 credito = 0,01 USD

Calcolatore di costi

Calcolatore prezzi

/ richiesta
/ richiesta

Totale

0,24 USD

24 crediti

Per richiesta

0,0024 USD · 0,24 crediti

Ogni richiesta viene arrotondata a 0,01 crediti.

04

API

Chiama Gemini 3 Flash con la tua chiave API Railwail. Usa questo ID modello nella richiesta:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-3-flash",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Imposta la tua chiave come RAILWAIL_API_KEYCrea chiave API
05

Specifiche

ID modello
gemini-3-flash
Sviluppatore
Google DeepMind
Categoria
Multimodale
Input
Testo, Immagine, Audio, Video
Output
Testo
Finestra di contesto
1.048.576 token
Output massimo
65.536 token
Fatturazione
In base all'utilizzo (token o tempo GPU)
Ciclo di vita
Deprecato
Dimensione del modello
Undisclosed (sparse MoE, smaller and sparser than Gemini 3.1 Pro)
Licenza
Proprietary commercial license via Google AI Studio, Vertex AI and the Gemini app. Free tier available in the Gemini app and AI Mode in Search.
Voce di catalogo aggiornata
23 settembre 2026

Parametri di input

Input e impostazioni dallo schema di input del modello. L'esempio nella sezione API mostra quali di essi l'API accetta.

  • promptObbligatorio

    User message

    Tipo: Testo
    Predefinito: –
    Valori consentiti: fino a 32.000 caratteri
  • top_p
    Tipo: Numero
    Predefinito: 0.95
    Valori consentiti: 0 a 1
  • stream
    Tipo: Sì/No
    Predefinito: false
    Valori consentiti: –
  • image_url

    Optional image URL to analyze

    Tipo: Testo
    Predefinito: –
    Valori consentiti: –
  • max_tokens
    Tipo: Numero intero
    Predefinito: 4096
    Valori consentiti: 1 a 32.000
  • temperature
    Tipo: Numero
    Predefinito: 1
    Valori consentiti: 0 a 2
  • system_prompt

    Optional system instruction

    Tipo: Testo
    Predefinito: –
    Valori consentiti: fino a 8000 caratteri

Etichette

  • google
  • deepmind
  • balanced
  • multimodal
  • low-latency
  • long-context
  • 1m-context
06

Casi d'uso

A cosa serve

  • Default consumer multimodal chat
  • High-throughput agentic workflows
  • Real-time RAG pipelines
  • Long-video summarisation and search
  • Production coding assistants
  • Voice and audio reasoning backends
  • AI Mode in Search and Antigravity workflows
07

Domande frequenti

Cos'è Gemini 3 Flash?

Gemini 3 Flash è un modello di Google DeepMind nella categoria Multimodale. Su Railwail puoi richiamarlo con una chiave API tramite l'API Railwail.

Quanto costa Gemini 3 Flash su Railwail?

Su Railwail, Gemini 3 Flash costa 0,60 USD per 1M token di input e 3,60 USD per 1M token di output. Ti viene addebitato ciò che ogni richiesta utilizza effettivamente. L'utilizzo viene pagato con crediti prepagati; 1 credito equivale a 0,01 USD.

Qual è la finestra di contesto di Gemini 3 Flash?

La finestra di contesto di Gemini 3 Flash contiene 1.048.576 token. Una risposta può essere lunga fino a 65.536 token.

Quanto è veloce Gemini 3 Flash?

Non ci sono ancora abbastanza esecuzioni misurate di Gemini 3 Flash su Railwail per indicare un tempo di esecuzione. Dipende dall'input, dalle impostazioni e dal carico presso il provider.

Gemini 3 Flash è migliore di BLIP?

Dipende dall'attività. Gemini 3 Flash (Google DeepMind) e BLIP (Salesforce) sono entrambi modelli nella categoria Multimodale. La pagina di confronto mostra i loro prezzi e le specifiche affiancati.

Confronta Gemini 3 Flash e BLIP

Gemini 3 Flash può elaborare immagini?

Sì. Gemini 3 Flash accetta immagini come input oltre al testo.

Come uso Gemini 3 Flash tramite l'API?

Crea una chiave API Railwail e invia la tua richiesta con l'ID modello gemini-3-flash. Gli esempi di codice per curl, Python e JavaScript sono nella sezione API di questa pagina.

08

Modelli comparabili

Tutti in questa categoria

Usa Gemini 3 Flash tramite l'API

Una chiave API per tutti i modelli su Railwail. L'utilizzo viene addebitato da crediti prepagati, 1 credito = 0,01 USD.