DeepSeek V4 Flash

Κείμενο & συνομιλίαΛηγμένοΔιαθέσιμο
από DeepSeekΑναγνωριστικό μοντέλου: deepseek-v4-flash

Efficiency-optimized variant of DeepSeek V4. 284B MoE / 13B active, 1M context, ultra-low pricing for high-throughput workloads.

Τιμή · 1M εισόδου / εξόδου
0,36 $ / 1,44 $
Περιεχόμενο
1.048.575 tokens
Μέγ. έξοδος
384.000 tokens
Είσοδος → Έξοδος
Κείμενο → Κείμενο
Χρόνος εκτέλεσης (διάμεσος)
1,3 s
Προγραμματιστής
DeepSeek

Ο πάροχος διακόπτει αυτό το μοντέλο.

Διαθέσιμη νεότερη έκδοση: DeepSeek V4.1 Flash

01

Playground

Δοκιμάστε DeepSeek V4 Flash

Συνομιλία

0,36 $/1M in
Δοκιμάστε DeepSeek V4 Flash

Στείλτε ένα μήνυμα. Η απάντηση φτάνει πλήρης όταν το μοντέλο τελειώσει (χωρίς streaming).

System prompt
Μέγ. μήκος απάντησης (tokens)

Αυτή η εκτέλεση

το πολύ 0,0015 $ · 0,15 πιστωτικές μονάδες προεπιλεγμένες

Χρεώνονται τα tokens που χρησιμοποιήθηκαν πραγματικά· το αχρησιμοποίητο μέρος της προεπιλογής επιστρέφεται.

Νέος εδώ;

10 δωρεάν πιστωτικές μονάδες (0,10 $) όταν εγγραφείς με Google

Χρησιμοποιήσιμο 24 ώρες μετά την εγγραφή, έως 5 εκτελέσεις ανά ημέρα και το πολύ 2 πιστωτικές μονάδες ανά εκτέλεση. Άλλες μέθοδοι σύνδεσης ξεκινούν χωρίς πιστωτικές μονάδες. Αρκετό για 66 εκτελέσεις αυτού του μοντέλου.

02

Σχετικά με το DeepSeek V4 Flash

ΣύντομαΗμερομηνία: 23 Σεπτεμβρίου 2026

Το DeepSeek V4 Flash είναι ένα μοντέλο του DeepSeek στην κατηγορία Κείμενο & συνομιλία. Στο Railwail, το DeepSeek V4 Flash κοστίζει 0,36 $ ανά 1M tokens εισόδου και 1,44 $ ανά 1M tokens εξόδου. Το παράθυρο περιεχομένου περιέχει 1.048.575 tokens, και μια απάντηση μπορεί να είναι έως 384.000 tokens. Νεότερη έκδοση: DeepSeek V4.1 Flash.

DeepSeek-V4-Flash is the cost-efficient sibling of V4-Pro, released April 2026 as part of the V4 Preview. 284B total / 13B active MoE parameters with the same 1M-token context window. Designed for high-throughput agentic loops, RAG and batch tasks where latency and cost matter more than raw capability. Recommended for production agents, classification at scale, large-scale data extraction.

Φόντο

Σχετικά με DeepSeek AI

Ιδρύθηκε 2023 · Hangzhou, China

DeepSeek AI is a Chinese AI research lab founded in 2023 by Liang Wenfeng, founder of the High-Flyer quantitative hedge fund. The lab is funded primarily by High-Flyer's profits. Its mission is open frontier AI, with all flagship models released with open weights. Major releases include DeepSeek LLM (2023), DeepSeek-V2 (May 2024), DeepSeek-V3 (December 2024), DeepSeek-R1 (January 2026), DeepSeek V3.1 (early 2026) and the DeepSeek V4 family (April 24, 2026), comprising V4-Pro and V4-Flash. DeepSeek is credited with popularising large-scale Reinforcement Learning from Verifiable Rewards and consistently tops open-weights leaderboards.

Επισκεφθείτε DeepSeek AI

Αρχιτεκτονική

Sparse Mixture-of-Experts Transformer (efficiency-optimized open-weights)

DeepSeek-V4-Flash was released April 24, 2026 as the efficiency-optimized sibling of V4-Pro. It is a Sparse MoE Transformer with 284B total parameters and 13B activated per token, retaining the full 1M-token native context window and 384K-token max output of the Pro variant at significantly lower inference cost. The model uses the same DeepSeek architectural stack: Multi-head Latent Attention (MLA), DeepSeekMoE with fine-grained expert specialization and shared experts, and FP8 mixed-precision training. Post-training combined supervised fine-tuning, RLVR on math/code/tool-use trajectories, and heavy distillation from the V4-Pro teacher model. V4 Flash is published with open weights under a permissive license and is designed for production-scale RAG, agentic loops and high-throughput workloads. At $0.112 input / $0.224 output per million tokens it undercuts every Western frontier model by an order of magnitude.

Παράμετροι
284B total / 13B active per token
Περιεχόμενο
1.048.575 tokens

Δυνατότητες

  • 1M token native context window with 384K max output
  • 284B MoE / 13B active parameters
  • Ultra-low pricing ($0.112 / $0.224 per million tokens)
  • Distilled from DeepSeek V4-Pro teacher model
  • FP8-trained for compute efficiency
  • Multi-head Latent Attention for memory-efficient long context
  • Function calling and structured JSON output
  • Strong on math, STEM and coding for its size
  • Available via DeepSeek API, OpenRouter, Together and self-hosted with vLLM/SGLang
  • Open weights under a permissive license
  • Best for: production agents, RAG pipelines, high-throughput data extraction, on-premise inference under tight cost budgets.

Εκπαίδευση & άδεια

Pretrained on the same multi-trillion-token mixture as V4-Pro. Post-training combines supervised fine-tuning, RLVR and distillation from the V4-Pro teacher model. Knowledge cutoff approximately early 2026.

Άδεια: Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.

Δοκιμές ασφάλειας: DeepSeek publishes model cards but provides limited external red-teaming. Safety filters are lighter than Western frontier labs; deployers are responsible for downstream alignment.

Γνωστοί περιορισμοί

  • Below V4-Pro on the hardest reasoning and coding benchmarks
  • Light built-in safety alignment relative to Western frontier models
  • No native vision or audio input (text-only)
  • Older deepseek-chat / deepseek-reasoner endpoints will be deprecated July 24, 2026
  • Some Chinese-language safety constraints apply
03

Τιμολόγηση

Οι τιμές είναι σε δολάρια ΗΠΑ. Η χρήση χρεώνεται από προπληρωμένα πιστωτικά.
Είσοδος0,36 $ / 1M tokens
Έξοδος1,44 $ / 1M tokens
  • Χρέωση με βάση τα tokens που χρησιμοποιεί πραγματικά κάθε αίτημα.
  • 1 πιστωτική μονάδα = 0,01 $

Αριθμομηχανή κόστους

Αριθμομηχανή τιμών

/ αίτημα
/ αίτημα

Σύνολο

0,11 $

11 πιστωτικές μονάδες

Ανά αίτημα

0,0011 $ · 0,11 πιστωτικές μονάδες

Κάθε αίτημα στρογγυλοποιείται προς τα πάνω σε 0,01 πιστωτικές μονάδες.

04

API

Καλέστε το DeepSeek V4 Flash με το κλειδί Railwail API. Χρησιμοποιήστε αυτό το ID μοντέλου στο αίτημα:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Ορίστε το κλειδί σας ως RAILWAIL_API_KEYΔημιουργία κλειδιού API
05

Προδιαγραφές

ID μοντέλου
deepseek-v4-flash
Ανάπτυξη
DeepSeek
Είσοδος
Κείμενο
Έξοδος
Κείμενο
Παράθυρο περιεχομένου
1.048.575 tokens
Μέγ. έξοδος
384.000 tokens
Χρέωση
Κατά χρήση (tokens ή χρόνος GPU)
Χρόνος εκτέλεσης (διάμεσος)
1,3 s15 ολοκληρωμένες εκτελέσεις στο Railwail τις τελευταίες 90 ημέρες
Κύκλος ζωής
Ληγμένο
Μέγεθος μοντέλου
284B total / 13B active per token
Άδεια
Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.
Καταχώρηση καταλόγου ενημερώθηκε
23 Σεπτεμβρίου 2026

Ετικέτες

  • deepseek
  • open-weights
  • moe
  • cost-efficient
  • long-context
  • 1m-context
06

Περιπτώσεις χρήσης

Για τι χρησιμοποιείται

  • Production RAG pipelines
  • High-throughput coding subagents
  • Bulk data extraction and classification
  • Cost-sensitive enterprise APIs
  • On-premise inference under tight cost budgets
  • Long-document summarisation at scale
  • Real-time chat backends
07

Συχνές ερωτήσεις

Τι είναι DeepSeek V4 Flash;

Το DeepSeek V4 Flash είναι ένα μοντέλο του DeepSeek στην κατηγορία Κείμενο & συνομιλία. Στο Railwail μπορείτε να το καλέσετε με ένα API key μέσω του Railwail API.

Πόσο κοστίζει το DeepSeek V4 Flash στο Railwail;

Στο Railwail, το DeepSeek V4 Flash κοστίζει 0,36 $ ανά 1M tokens εισόδου και 1,44 $ ανά 1M tokens εξόδου. Χρεώνεστε για αυτό που χρησιμοποιεί πραγματικά κάθε αίτημα. Η χρήση πληρώνεται από προπληρωμένες πιστωτικές μονάδες· 1 πιστωτική μονάδα ισούται με 0,01 $.

Ποιο είναι το παράθυρο περιεχομένου του DeepSeek V4 Flash;

Το παράθυρο περιεχομένου του DeepSeek V4 Flash περιέχει 1.048.575 tokens. Μια απάντηση μπορεί να είναι έως 384.000 tokens.

Πόσο γρήγορο είναι το DeepSeek V4 Flash;

Στο Railwail, ο διάμεσος χρόνος εκτέλεσης του DeepSeek V4 Flash τις τελευταίες 90 ημέρες ήταν 1,3 s, βάσει 15 ολοκληρωμένων εκτελέσεων.

Είναι το DeepSeek V4 Flash καλύτερο από το DeepSeek V4.1 Flash;

Αυτό εξαρτάται από την εργασία. Το DeepSeek V4 Flash (DeepSeek) και το DeepSeek V4.1 Flash (DeepSeek) είναι και τα δύο μοντέλα στην κατηγορία Κείμενο & συνομιλία. Η σελίδα σύγκρισης δείχνει τις τιμές και τις προδιαγραφές τους παράλληλα.

Σύγκριση DeepSeek V4 Flash και DeepSeek V4.1 Flash

Πώς χρησιμοποιώ το DeepSeek V4 Flash μέσω του API;

Δημιουργήστε ένα Railwail API key και στείλτε το αίτημά σας με το ID μοντέλου deepseek-v4-flash. Παραδείγματα κώδικα για curl, Python και JavaScript βρίσκονται στην ενότητα API αυτής της σελίδας.

08

Συγκρίσιμα μοντέλα

Όλα σε αυτήν την κατηγορία

Χρησιμοποιήστε το DeepSeek V4 Flash μέσω του API

Ένα κλειδί API για κάθε μοντέλο στο Railwail. Η χρήση χρεώνεται από προπληρωμένα πιστωτικά, 1 πιστωτικό = 0,01 $.