DeepSeek V4 Flash

Tekst i chatWycofywaneDostępne
od DeepSeekID modelu: deepseek-v4-flash

Efficiency-optimized variant of DeepSeek V4. 284B MoE / 13B active, 1M context, ultra-low pricing for high-throughput workloads.

Cena · 1M wejścia/wyjścia
0,36 USD / 1,44 USD
Kontekst
1 048 575 tokenów
Maks. wyjście
384 000 tokenów
Wejście → Wyjście
Tekst → Tekst
Czas wykonania (mediana)
1,3 s
Deweloper
DeepSeek

Dostawca wycofuje ten model.

Dostępna nowsza wersja: DeepSeek V4.1 Flash

01

Playground

Spróbuj DeepSeek V4 Flash

Chat

0,36 USD/1M wej.
Spróbuj DeepSeek V4 Flash

Wyślij wiadomość. Odpowiedź pojawi się w całości, gdy model będzie gotowy (bez streamingu).

Prompt systemowy
Maks. długość odpowiedzi (tokeny)

To uruchomienie

maksymalnie 0,0015 USD · 0,15 kredytów zarezerwowanych

Rozliczane są rzeczywiście użyte tokeny; niewykorzystana część rezerwacji jest zwracana.

Nowy tutaj?

10 darmowych kredytów (0,10 USD) po zarejestrowaniu się przez Google

Dostępne 24 godzin po rejestracji, do 5 uruchomień dziennie i maksymalnie 2 kredytów na uruchomienie. Inne metody logowania uruchamiają się bez kredytów. Wystarczy na 66 uruchomień tego modelu.

02

O DeepSeek V4 Flash

Krótko mówiącStan na 23 września 2026

DeepSeek V4 Flash to model opracowany przez DeepSeek w kategorii Tekst i chat. W serwisie Railwail DeepSeek V4 Flash kosztuje 0,36 USD za 1M tokenów wejściowych i 1,44 USD za 1M tokenów wyjściowych. Okno kontekstu zawiera 1 048 575 tokenów, a jedna odpowiedź może mieć do 384 000 tokenów. Nowsza wersja: DeepSeek V4.1 Flash.

DeepSeek-V4-Flash is the cost-efficient sibling of V4-Pro, released April 2026 as part of the V4 Preview. 284B total / 13B active MoE parameters with the same 1M-token context window. Designed for high-throughput agentic loops, RAG and batch tasks where latency and cost matter more than raw capability. Recommended for production agents, classification at scale, large-scale data extraction.

Tło

O DeepSeek AI

Założona 2023 · Hangzhou, China

DeepSeek AI is a Chinese AI research lab founded in 2023 by Liang Wenfeng, founder of the High-Flyer quantitative hedge fund. The lab is funded primarily by High-Flyer's profits. Its mission is open frontier AI, with all flagship models released with open weights. Major releases include DeepSeek LLM (2023), DeepSeek-V2 (May 2024), DeepSeek-V3 (December 2024), DeepSeek-R1 (January 2026), DeepSeek V3.1 (early 2026) and the DeepSeek V4 family (April 24, 2026), comprising V4-Pro and V4-Flash. DeepSeek is credited with popularising large-scale Reinforcement Learning from Verifiable Rewards and consistently tops open-weights leaderboards.

Odwiedź DeepSeek AI

Architektura

Sparse Mixture-of-Experts Transformer (efficiency-optimized open-weights)

DeepSeek-V4-Flash was released April 24, 2026 as the efficiency-optimized sibling of V4-Pro. It is a Sparse MoE Transformer with 284B total parameters and 13B activated per token, retaining the full 1M-token native context window and 384K-token max output of the Pro variant at significantly lower inference cost. The model uses the same DeepSeek architectural stack: Multi-head Latent Attention (MLA), DeepSeekMoE with fine-grained expert specialization and shared experts, and FP8 mixed-precision training. Post-training combined supervised fine-tuning, RLVR on math/code/tool-use trajectories, and heavy distillation from the V4-Pro teacher model. V4 Flash is published with open weights under a permissive license and is designed for production-scale RAG, agentic loops and high-throughput workloads. At $0.112 input / $0.224 output per million tokens it undercuts every Western frontier model by an order of magnitude.

Parametry
284B total / 13B active per token
Kontekst
1 048 575 tokenów

Możliwości

  • 1M token native context window with 384K max output
  • 284B MoE / 13B active parameters
  • Ultra-low pricing ($0.112 / $0.224 per million tokens)
  • Distilled from DeepSeek V4-Pro teacher model
  • FP8-trained for compute efficiency
  • Multi-head Latent Attention for memory-efficient long context
  • Function calling and structured JSON output
  • Strong on math, STEM and coding for its size
  • Available via DeepSeek API, OpenRouter, Together and self-hosted with vLLM/SGLang
  • Open weights under a permissive license
  • Best for: production agents, RAG pipelines, high-throughput data extraction, on-premise inference under tight cost budgets.

Trening i licencja

Pretrained on the same multi-trillion-token mixture as V4-Pro. Post-training combines supervised fine-tuning, RLVR and distillation from the V4-Pro teacher model. Knowledge cutoff approximately early 2026.

Licencja: Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.

Testy bezpieczeństwa: DeepSeek publishes model cards but provides limited external red-teaming. Safety filters are lighter than Western frontier labs; deployers are responsible for downstream alignment.

Znane ograniczenia

  • Below V4-Pro on the hardest reasoning and coding benchmarks
  • Light built-in safety alignment relative to Western frontier models
  • No native vision or audio input (text-only)
  • Older deepseek-chat / deepseek-reasoner endpoints will be deprecated July 24, 2026
  • Some Chinese-language safety constraints apply
03

Ceny

Ceny w dolarach amerykańskich. Użycie jest rozliczane z prepłaconych kredytów.
Wejście0,36 USD / 1M tokenów
Wyjście1,44 USD / 1M tokenów
  • Rozliczane są tokeny, które faktycznie zużywa każde żądanie.
  • 1 kredyt = 0,01 USD

Kalkulator kosztów

Kalkulator cen

/ żądanie
/ żądanie

Razem

0,11 USD

11 kredytów

Za żądanie

0,0011 USD · 0,11 kredytów

Każde żądanie jest zaokrąglane w górę do 0,01 kredytu.

04

API

Wywołaj DeepSeek V4 Flash za pomocą klucza API Railwail. Użyj tego ID modelu w żądaniu:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Ustaw klucz jako RAILWAIL_API_KEYUtwórz klucz API
05

Specyfikacje

ID modelu
deepseek-v4-flash
Deweloper
DeepSeek
Kategoria
Tekst i chat
Wejście
Tekst
Wyjście
Tekst
Okno kontekstu
1 048 575 tokenów
Maks. wyjście
384 000 tokenów
Rozliczenie
Według użycia (tokeny lub czas GPU)
Czas wykonania (mediana)
1,3 s15 ukończonych uruchomień na Railwail w ciągu ostatnich 90 dni
Cykl życia
Wycofywane
Rozmiar modelu
284B total / 13B active per token
Licencja
Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.
Wpis w katalogu zaktualizowany
23 września 2026

Tagi

  • deepseek
  • open-weights
  • moe
  • cost-efficient
  • long-context
  • 1m-context
06

Przypadki użycia

Do czego się go używa

  • Production RAG pipelines
  • High-throughput coding subagents
  • Bulk data extraction and classification
  • Cost-sensitive enterprise APIs
  • On-premise inference under tight cost budgets
  • Long-document summarisation at scale
  • Real-time chat backends
07

Często zadawane pytania

Co to jest DeepSeek V4 Flash?

DeepSeek V4 Flash to model opracowany przez DeepSeek w kategorii Tekst i chat. W serwisie Railwail możesz go wywołać za pomocą klucza API poprzez API Railwail.

Ile kosztuje DeepSeek V4 Flash w serwisie Railwail?

W serwisie Railwail DeepSeek V4 Flash kosztuje 0,36 USD za 1M tokenów wejściowych i 1,44 USD za 1M tokenów wyjściowych. Opłata jest pobierana za to, co faktycznie zużywa każde żądanie. Użycie jest opłacane z przedpłaconych kredytów; 1 kredyt równa się 0,01 USD.

Jakie jest okno kontekstu DeepSeek V4 Flash?

Okno kontekstu DeepSeek V4 Flash zawiera 1 048 575 tokenów. Jedna odpowiedź może mieć do 384 000 tokenów.

Jak szybki jest DeepSeek V4 Flash?

W serwisie Railwail mediana czasu przebiegu DeepSeek V4 Flash w ciągu ostatnich 90 dni wyniosła 1,3 s, na podstawie 15 ukończonych przebiegów.

Czy DeepSeek V4 Flash jest lepszy niż DeepSeek V4.1 Flash?

To zależy od zadania. DeepSeek V4 Flash (DeepSeek) i DeepSeek V4.1 Flash (DeepSeek) to oba modele z kategorii Tekst i chat. Strona porównania pokazuje ich ceny i specyfikacje obok siebie.

Porównaj DeepSeek V4 Flash i DeepSeek V4.1 Flash

Jak używać DeepSeek V4 Flash przez API?

Utwórz klucz API Railwail i wyślij swoje żądanie z ID modelu deepseek-v4-flash. Przykłady kodu dla curl, Python i JavaScript znajdują się w sekcji API na tej stronie.

08

Porównywalne modele

Wszystkie w tej kategorii

Użyj DeepSeek V4 Flash przez API

Jeden klucz API dla każdego modelu na Railwail. Opłaty pobierane są z przedpłaconych kredytów, 1 kredyt = 0,01 USD.