DeepSeek V4 Flash

Teksti & chatPoistuvaSaatavilla
kehittäjä: DeepSeekMallin tunnus: deepseek-v4-flash

Efficiency-optimized variant of DeepSeek V4. 284B MoE / 13B active, 1M context, ultra-low pricing for high-throughput workloads.

Hinta · 1M sisään/ulos
0,36 $ / 1,44 $
Konteksti
1 048 575 tokenia
Enint. tuloste
384 000 tokenia
Syöte → tulos
Teksti → Teksti
Suoritusaika (mediaani)
1,3 s
Kehittäjä
DeepSeek

Palveluntarjoaja poistaa tämän mallin käytöstä.

Uudempi versio saatavilla: DeepSeek V4.1 Flash

01

Leikkikenttä

Kokeile DeepSeek V4 Flash

Chat

0,36 $/1M in
Kokeile DeepSeek V4 Flash

Lähetä viesti. Vastaus saapuu kokonaisuudessaan, kun malli on valmis (ei suoratoistoa).

Järjestelmäkehote
Enint. vastausten pituus (tokenit)

Tämä suoritus

enintään 0,0015 $ · 0,15 creditiä varattu

Laskutus todella käytettyjen tokenien mukaan; varauksen käyttämätön osa palautetaan.

Uusi täällä?

10 ilmaista creditiä (0,10 $) kun rekisteröidyt Googlella

Käytettävissä 24 tuntia rekisteröinnin jälkeen, enintään 5 suoritusta päivässä ja enintään 2 creditiä suoritusta kohti. Muut kirjautumismenetelmät alkavat ilman creditejä. Riittää 66 suoritukseen tästä mallista.

02

Tietoja: DeepSeek V4 Flash

Lyhyesti23. syyskuuta 2026 alkaen

DeepSeek V4 Flash on DeepSeek-kehittäjän malli kategoriasta Teksti & chat. Railwailissa DeepSeek V4 Flash maksaa 0,36 $ per 1M syötetokenia ja 1,44 $ per 1M tulostetokenia. Kontekstiikkuna sisältää 1 048 575 tokenia, ja vastaus voi olla enintään 384 000 tokenia pitkä. Uudempi versio: DeepSeek V4.1 Flash.

DeepSeek-V4-Flash is the cost-efficient sibling of V4-Pro, released April 2026 as part of the V4 Preview. 284B total / 13B active MoE parameters with the same 1M-token context window. Designed for high-throughput agentic loops, RAG and batch tasks where latency and cost matter more than raw capability. Recommended for production agents, classification at scale, large-scale data extraction.

Tausta

Tietoja: DeepSeek AI

Perustettu 2023 · Hangzhou, China

DeepSeek AI is a Chinese AI research lab founded in 2023 by Liang Wenfeng, founder of the High-Flyer quantitative hedge fund. The lab is funded primarily by High-Flyer's profits. Its mission is open frontier AI, with all flagship models released with open weights. Major releases include DeepSeek LLM (2023), DeepSeek-V2 (May 2024), DeepSeek-V3 (December 2024), DeepSeek-R1 (January 2026), DeepSeek V3.1 (early 2026) and the DeepSeek V4 family (April 24, 2026), comprising V4-Pro and V4-Flash. DeepSeek is credited with popularising large-scale Reinforcement Learning from Verifiable Rewards and consistently tops open-weights leaderboards.

Vieraile sivustolla DeepSeek AI

Arkkitehtuuri

Sparse Mixture-of-Experts Transformer (efficiency-optimized open-weights)

DeepSeek-V4-Flash was released April 24, 2026 as the efficiency-optimized sibling of V4-Pro. It is a Sparse MoE Transformer with 284B total parameters and 13B activated per token, retaining the full 1M-token native context window and 384K-token max output of the Pro variant at significantly lower inference cost. The model uses the same DeepSeek architectural stack: Multi-head Latent Attention (MLA), DeepSeekMoE with fine-grained expert specialization and shared experts, and FP8 mixed-precision training. Post-training combined supervised fine-tuning, RLVR on math/code/tool-use trajectories, and heavy distillation from the V4-Pro teacher model. V4 Flash is published with open weights under a permissive license and is designed for production-scale RAG, agentic loops and high-throughput workloads. At $0.112 input / $0.224 output per million tokens it undercuts every Western frontier model by an order of magnitude.

Parametrit
284B total / 13B active per token
Konteksti
1 048 575 tokenia

Ominaisuudet

  • 1M token native context window with 384K max output
  • 284B MoE / 13B active parameters
  • Ultra-low pricing ($0.112 / $0.224 per million tokens)
  • Distilled from DeepSeek V4-Pro teacher model
  • FP8-trained for compute efficiency
  • Multi-head Latent Attention for memory-efficient long context
  • Function calling and structured JSON output
  • Strong on math, STEM and coding for its size
  • Available via DeepSeek API, OpenRouter, Together and self-hosted with vLLM/SGLang
  • Open weights under a permissive license
  • Best for: production agents, RAG pipelines, high-throughput data extraction, on-premise inference under tight cost budgets.

Koulutus ja lisenssi

Pretrained on the same multi-trillion-token mixture as V4-Pro. Post-training combines supervised fine-tuning, RLVR and distillation from the V4-Pro teacher model. Knowledge cutoff approximately early 2026.

Lisenssi: Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.

Turvallisuustestit: DeepSeek publishes model cards but provides limited external red-teaming. Safety filters are lighter than Western frontier labs; deployers are responsible for downstream alignment.

Tunnetut rajoitukset

  • Below V4-Pro on the hardest reasoning and coding benchmarks
  • Light built-in safety alignment relative to Western frontier models
  • No native vision or audio input (text-only)
  • Older deepseek-chat / deepseek-reasoner endpoints will be deprecated July 24, 2026
  • Some Chinese-language safety constraints apply
03

Hinnat

Hinnat Yhdysvaltain dollareissa. Käyttö laskutetaan prepaid-krediiteistä.
Syöte0,36 $ / 1M tokenia
Tuloste1,44 $ / 1M tokenia
  • Laskutus perustuu kunkin pyynnön todella käyttämiin tokeneihin.
  • 1 krediitti = 0,01 $

Kustannuslaskin

Hintalaskin

/ pyynnöllä
/ pyynnöllä

Yhteensä

0,11 $

11 krediittiä

Pyynnöllä

0,0011 $ · 0,11 krediittiä

Jokainen pyyntö pyöristetään ylöspäin 0,01 krediittiin.

04

API

Kutsu DeepSeek V4 Flash Railwail-API-avaimellasi. Käytä tätä mallin tunnusta pyynnössä:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Aseta avaimesi muuttujaksi RAILWAIL_API_KEYLuo API-avain
05

Tekniset tiedot

Mallin tunnus
deepseek-v4-flash
Kehittäjä
DeepSeek
Kategoria
Teksti & chat
Syöte
Teksti
Tuloste
Teksti
Konteksti-ikkuna
1 048 575 tokenia
Enimmäistuloste
384 000 tokenia
Laskutus
Käytön mukaan (tokeneja tai GPU-aikaa)
Suoritusaika (mediaani)
1,3 s15 suoritettua ajoa Railwailissa viimeisten 90 päivän aikana
Elinkaari
Poistuva
Mallin koko
284B total / 13B active per token
Lisenssi
Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.
Luettelokirjaus päivitetty
23. syyskuuta 2026

Tunnisteet

  • deepseek
  • open-weights
  • moe
  • cost-efficient
  • long-context
  • 1m-context
06

Käyttötapaukset

Mihin sitä käytetään

  • Production RAG pipelines
  • High-throughput coding subagents
  • Bulk data extraction and classification
  • Cost-sensitive enterprise APIs
  • On-premise inference under tight cost budgets
  • Long-document summarisation at scale
  • Real-time chat backends
07

Usein kysytyt kysymykset

Mikä on DeepSeek V4 Flash?

DeepSeek V4 Flash on DeepSeekn kehittämä malli Teksti & chat-kategoriassa. Railwailissa voit kutsua sitä API-avaimella Railwail-ohjelmointirajapinnan kautta.

Paljonko DeepSeek V4 Flash maksaa Railwailissa?

Railwailissa DeepSeek V4 Flash maksaa 0,36 $ per 1M syötetokenia ja 1,44 $ per 1M tulostetokenia. Sinua veloitetaan siitä, mitä kukin pyyntö todella käyttää. Käyttö maksetaan ennakkoon ostettujen krediittien avulla; 1 krediitti vastaa 0,01 $.

Mikä on DeepSeek V4 Flashn kontekstiikkuna?

DeepSeek V4 Flashn kontekstiikkuna sisältää 1 048 575 tokenia. Vastaus voi olla enintään 384 000 tokenia pitkä.

Kuinka nopea DeepSeek V4 Flash on?

Railwailissa DeepSeek V4 Flashn mediaanisuoritusaika viimeisten 90 päivän aikana oli 1,3 s, perustuen 15 suoritettuun suoritukseen.

Onko DeepSeek V4 Flash parempi kuin DeepSeek V4.1 Flash?

Se riippuu tehtävästä. DeepSeek V4 Flash (DeepSeek) ja DeepSeek V4.1 Flash (DeepSeek) ovat molemmat malleja Teksti & chat-kategoriassa. Vertailussa näkyvät niiden hinnat ja tekniset tiedot rinnakkain.

Vertaa DeepSeek V4 Flash ja DeepSeek V4.1 Flash

Kuinka käytän DeepSeek V4 Flasha ohjelmointirajapinnan kautta?

Luo Railwail-ohjelmointirajapinnan avain ja lähetä pyyntösi mallin tunnuksella deepseek-v4-flash. Koodiesimerkit curlille, Pythonille ja JavaScriptille ovat tämän sivun API-osiossa.

08

Vertailukelpoiset mallit

Kaikki tässä kategoriassa

Käytä DeepSeek V4 Flash API:n kautta

Yksi API-avain kaikille Railwailin malleille. Käyttö laskutetaan prepaid-krediiteistä, 1 krediitti = 0,01 $.