DeepSeek V4 Flash

Tekst & chatAfgelopenBeschikbaar
van DeepSeekModel-ID: deepseek-v4-flash

Efficiency-optimized variant of DeepSeek V4. 284B MoE / 13B active, 1M context, ultra-low pricing for high-throughput workloads.

Prijs · 1M in/uit
US$ 0,36 / US$ 1,44
Context
1.048.575 tokens
Max. uitvoer
384.000 tokens
Invoer → Uitvoer
Tekst → Tekst
Uitvoeringstijd (mediaan)
1,3 s
Ontwikkelaar
DeepSeek

De provider faset dit model uit.

Nieuwere versie beschikbaar: DeepSeek V4.1 Flash

01

Playground

DeepSeek V4 Flash proberen

Chat

US$ 0,36/1M in
DeepSeek V4 Flash proberen

Stuur een bericht. Het antwoord komt volledig binnen zodra het model klaar is (geen streaming).

Systeemprompt
Max. antwoordlengte (tokens)

Deze uitvoering

maximaal US$ 0,0015 · 0,15 credits gereserveerd

Gefactureerd worden de werkelijk gebruikte tokens; het ongebruikte deel van de reservering wordt terugbetaald.

Nieuw hier?

10 gratis credits (US$ 0,10) wanneer je je aanmeldt met Google

Bruikbaar 24 uur na aanmelding, tot 5 uitvoeringen per dag en maximaal 2 credits per uitvoering. Andere aanmeldmethoden starten zonder credits. Voldoende voor 66 uitvoeringen van dit model.

02

Over DeepSeek V4 Flash

SamengevatPer 23 september 2026

DeepSeek V4 Flash is een model van DeepSeek in de categorie Tekst & chat. Op Railwail kost DeepSeek V4 Flash US$ 0,36 per 1M invoertokens en US$ 1,44 per 1M uitvoertokens. Het contextvenster bevat 1.048.575 tokens, en een antwoord kan tot 384.000 tokens lang zijn. Nieuwere versie: DeepSeek V4.1 Flash.

DeepSeek-V4-Flash is the cost-efficient sibling of V4-Pro, released April 2026 as part of the V4 Preview. 284B total / 13B active MoE parameters with the same 1M-token context window. Designed for high-throughput agentic loops, RAG and batch tasks where latency and cost matter more than raw capability. Recommended for production agents, classification at scale, large-scale data extraction.

Achtergrond

Over DeepSeek AI

Opgericht 2023 · Hangzhou, China

DeepSeek AI is a Chinese AI research lab founded in 2023 by Liang Wenfeng, founder of the High-Flyer quantitative hedge fund. The lab is funded primarily by High-Flyer's profits. Its mission is open frontier AI, with all flagship models released with open weights. Major releases include DeepSeek LLM (2023), DeepSeek-V2 (May 2024), DeepSeek-V3 (December 2024), DeepSeek-R1 (January 2026), DeepSeek V3.1 (early 2026) and the DeepSeek V4 family (April 24, 2026), comprising V4-Pro and V4-Flash. DeepSeek is credited with popularising large-scale Reinforcement Learning from Verifiable Rewards and consistently tops open-weights leaderboards.

DeepSeek AI bezoeken

Architectuur

Sparse Mixture-of-Experts Transformer (efficiency-optimized open-weights)

DeepSeek-V4-Flash was released April 24, 2026 as the efficiency-optimized sibling of V4-Pro. It is a Sparse MoE Transformer with 284B total parameters and 13B activated per token, retaining the full 1M-token native context window and 384K-token max output of the Pro variant at significantly lower inference cost. The model uses the same DeepSeek architectural stack: Multi-head Latent Attention (MLA), DeepSeekMoE with fine-grained expert specialization and shared experts, and FP8 mixed-precision training. Post-training combined supervised fine-tuning, RLVR on math/code/tool-use trajectories, and heavy distillation from the V4-Pro teacher model. V4 Flash is published with open weights under a permissive license and is designed for production-scale RAG, agentic loops and high-throughput workloads. At $0.112 input / $0.224 output per million tokens it undercuts every Western frontier model by an order of magnitude.

Parameters
284B total / 13B active per token
Context
1.048.575 tokens

Mogelijkheden

  • 1M token native context window with 384K max output
  • 284B MoE / 13B active parameters
  • Ultra-low pricing ($0.112 / $0.224 per million tokens)
  • Distilled from DeepSeek V4-Pro teacher model
  • FP8-trained for compute efficiency
  • Multi-head Latent Attention for memory-efficient long context
  • Function calling and structured JSON output
  • Strong on math, STEM and coding for its size
  • Available via DeepSeek API, OpenRouter, Together and self-hosted with vLLM/SGLang
  • Open weights under a permissive license
  • Best for: production agents, RAG pipelines, high-throughput data extraction, on-premise inference under tight cost budgets.

Training & licentie

Pretrained on the same multi-trillion-token mixture as V4-Pro. Post-training combines supervised fine-tuning, RLVR and distillation from the V4-Pro teacher model. Knowledge cutoff approximately early 2026.

Licentie: Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.

Veiligheidstests: DeepSeek publishes model cards but provides limited external red-teaming. Safety filters are lighter than Western frontier labs; deployers are responsible for downstream alignment.

Bekende beperkingen

  • Below V4-Pro on the hardest reasoning and coding benchmarks
  • Light built-in safety alignment relative to Western frontier models
  • No native vision or audio input (text-only)
  • Older deepseek-chat / deepseek-reasoner endpoints will be deprecated July 24, 2026
  • Some Chinese-language safety constraints apply
03

Prijzen

Prijzen in US-dollars. Het gebruik wordt in rekening gebracht via vooraf gekochte credits.
InvoerUS$ 0,36 / 1M tokens
UitvoerUS$ 1,44 / 1M tokens
  • Gefactureerd worden de tokens die elke aanvraag daadwerkelijk gebruikt.
  • 1 credit = US$ 0,01

Kostencalculator

Prijscalculator

/ aanvr.
/ aanvr.

Totaal

US$ 0,11

11 credits

Per aanvraag

US$ 0,0011 · 0,11 credits

Elke aanvraag wordt afgerond naar 0,01 credits.

04

API

Roep DeepSeek V4 Flash aan met je Railwail API-sleutel. Gebruik deze model-ID in het verzoek:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Stel uw sleutel in als RAILWAIL_API_KEYAPI-sleutel maken
05

Specificaties

Model-ID
deepseek-v4-flash
Ontwikkelaar
DeepSeek
Categorie
Tekst & chat
Invoer
Tekst
Uitvoer
Tekst
Contextvenster
1.048.575 tokens
Max. uitvoer
384.000 tokens
Facturering
Op basis van gebruik (tokens of GPU-tijd)
Uitvoeringstijd (mediaan)
1,3 s15 voltooide uitvoeringen op Railwail in de afgelopen 90 dagen
Levenscyclus
Afgelopen
Modelgrootte
284B total / 13B active per token
Licentie
Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.
Catalogusitem bijgewerkt
23 september 2026

Tags

  • deepseek
  • open-weights
  • moe
  • cost-efficient
  • long-context
  • 1m-context
06

Gebruiksscenario's

Waarvoor het wordt gebruikt

  • Production RAG pipelines
  • High-throughput coding subagents
  • Bulk data extraction and classification
  • Cost-sensitive enterprise APIs
  • On-premise inference under tight cost budgets
  • Long-document summarisation at scale
  • Real-time chat backends
07

Veelgestelde vragen

Wat is DeepSeek V4 Flash?

DeepSeek V4 Flash is een model van DeepSeek in de categorie Tekst & chat. Op Railwail kunt u het aanroepen met een API-sleutel via de Railwail-API.

Hoeveel kost DeepSeek V4 Flash op Railwail?

Op Railwail kost DeepSeek V4 Flash US$ 0,36 per 1M invoertokens en US$ 1,44 per 1M uitvoertokens. U betaalt voor wat elke aanvraag daadwerkelijk verbruikt. Gebruik wordt betaald met vooraf gekochte credits; 1 credit is gelijk aan US$ 0,01.

Wat is het contextvenster van DeepSeek V4 Flash?

Het contextvenster van DeepSeek V4 Flash bevat 1.048.575 tokens. Een antwoord kan tot 384.000 tokens lang zijn.

Hoe snel is DeepSeek V4 Flash?

Op Railwail bedroeg de mediane uitvoeringstijd van DeepSeek V4 Flash in de afgelopen 90 dagen 1,3 s, gebaseerd op 15 voltooide runs.

Is DeepSeek V4 Flash beter dan DeepSeek V4.1 Flash?

Dat hangt van de taak af. DeepSeek V4 Flash (DeepSeek) en DeepSeek V4.1 Flash (DeepSeek) zijn beide modellen in de categorie Tekst & chat. De vergelijkingspagina toont hun prijzen en specificaties naast elkaar.

DeepSeek V4 Flash en DeepSeek V4.1 Flash vergelijken

Hoe gebruik ik DeepSeek V4 Flash via de API?

Maak een Railwail-API-sleutel aan en stuur uw aanvraag met de model-ID deepseek-v4-flash. Codevoorbeelden voor curl, Python en JavaScript staan in het API-gedeelte van deze pagina.

08

Vergelijkbare modellen

Alle in deze categorie

DeepSeek V4 Flash via de API gebruiken

Één API-sleutel voor elk model op Railwail. Gebruik wordt afgerekend via vooraf gekochte credits, 1 credit = US$ 0,01.