DeepSeek V4 Flash

Текст и чатСнято с производстваДоступно
от DeepSeekID модели: deepseek-v4-flash

Efficiency-optimized variant of DeepSeek V4. 284B MoE / 13B active, 1M context, ultra-low pricing for high-throughput workloads.

Цена · 1M вход/выход
0,36 $ / 1,44 $
Контекст
1 048 575 токенов
Макс. выход
384 000 токенов
Вход → выход
Текст → Текст
Время выполнения (медиана)
1,3 s
Разработчик
DeepSeek

Поставщик снимает эту модель с производства.

Доступна новая версия: DeepSeek V4.1 Flash

01

Playground

Попробовать DeepSeek V4 Flash

Чат

0,36 $/1M вход
Попробовать DeepSeek V4 Flash

Отправьте сообщение. Ответ придёт полностью, когда модель закончит работу (без потоковой передачи).

Системный промпт
Макс. длина ответа (токены)

Этот запуск

максимум 0,0015 $ · 0,15 кредитов зарезервировано

Оплачиваются только использованные токены; неиспользованная часть резервирования возвращается.

Впервые здесь?

10 бесплатных кредитов (0,10 $) при регистрации через Google

Доступно 24 часов после регистрации, до 5 запусков в день и максимум 2 кредитов за запуск. Другие способы входа начинают без кредитов. Достаточно для 66 запусков этой модели.

02

О DeepSeek V4 Flash

КороткоПо состоянию на 23 сентября 2026 г.

DeepSeek V4 Flash — это модель от DeepSeek в категории Текст и чат. На Railwail DeepSeek V4 Flash стоит 0,36 $ за 1M входных токенов и 1,44 $ за 1M выходных токенов. Контекстное окно содержит 1 048 575 токенов, ответ может быть до 384 000 токенов. Новая версия: DeepSeek V4.1 Flash.

DeepSeek-V4-Flash is the cost-efficient sibling of V4-Pro, released April 2026 as part of the V4 Preview. 284B total / 13B active MoE parameters with the same 1M-token context window. Designed for high-throughput agentic loops, RAG and batch tasks where latency and cost matter more than raw capability. Recommended for production agents, classification at scale, large-scale data extraction.

Фон

О DeepSeek AI

Основана в 2023 · Hangzhou, China

DeepSeek AI is a Chinese AI research lab founded in 2023 by Liang Wenfeng, founder of the High-Flyer quantitative hedge fund. The lab is funded primarily by High-Flyer's profits. Its mission is open frontier AI, with all flagship models released with open weights. Major releases include DeepSeek LLM (2023), DeepSeek-V2 (May 2024), DeepSeek-V3 (December 2024), DeepSeek-R1 (January 2026), DeepSeek V3.1 (early 2026) and the DeepSeek V4 family (April 24, 2026), comprising V4-Pro and V4-Flash. DeepSeek is credited with popularising large-scale Reinforcement Learning from Verifiable Rewards and consistently tops open-weights leaderboards.

Посетить DeepSeek AI

Архитектура

Sparse Mixture-of-Experts Transformer (efficiency-optimized open-weights)

DeepSeek-V4-Flash was released April 24, 2026 as the efficiency-optimized sibling of V4-Pro. It is a Sparse MoE Transformer with 284B total parameters and 13B activated per token, retaining the full 1M-token native context window and 384K-token max output of the Pro variant at significantly lower inference cost. The model uses the same DeepSeek architectural stack: Multi-head Latent Attention (MLA), DeepSeekMoE with fine-grained expert specialization and shared experts, and FP8 mixed-precision training. Post-training combined supervised fine-tuning, RLVR on math/code/tool-use trajectories, and heavy distillation from the V4-Pro teacher model. V4 Flash is published with open weights under a permissive license and is designed for production-scale RAG, agentic loops and high-throughput workloads. At $0.112 input / $0.224 output per million tokens it undercuts every Western frontier model by an order of magnitude.

Параметры
284B total / 13B active per token
Контекст
1 048 575 токенов

Возможности

  • 1M token native context window with 384K max output
  • 284B MoE / 13B active parameters
  • Ultra-low pricing ($0.112 / $0.224 per million tokens)
  • Distilled from DeepSeek V4-Pro teacher model
  • FP8-trained for compute efficiency
  • Multi-head Latent Attention for memory-efficient long context
  • Function calling and structured JSON output
  • Strong on math, STEM and coding for its size
  • Available via DeepSeek API, OpenRouter, Together and self-hosted with vLLM/SGLang
  • Open weights under a permissive license
  • Best for: production agents, RAG pipelines, high-throughput data extraction, on-premise inference under tight cost budgets.

Обучение и лицензия

Pretrained on the same multi-trillion-token mixture as V4-Pro. Post-training combines supervised fine-tuning, RLVR and distillation from the V4-Pro teacher model. Knowledge cutoff approximately early 2026.

Лицензия: Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.

Тестирование безопасности: DeepSeek publishes model cards but provides limited external red-teaming. Safety filters are lighter than Western frontier labs; deployers are responsible for downstream alignment.

Известные ограничения

  • Below V4-Pro on the hardest reasoning and coding benchmarks
  • Light built-in safety alignment relative to Western frontier models
  • No native vision or audio input (text-only)
  • Older deepseek-chat / deepseek-reasoner endpoints will be deprecated July 24, 2026
  • Some Chinese-language safety constraints apply
03

Цены

Цены в долларах США. Использование оплачивается из предоплаченных кредитов.
Входные данные0,36 $ / 1M токенов
Выходные данные1,44 $ / 1M токенов
  • Оплата рассчитывается по токенам, которые фактически использует каждый запрос.
  • 1 кредит = 0,01 $

Калькулятор стоимости

Калькулятор цен

/ запр.
/ запр.

Итого

0,11 $

11 кредитов

За запрос

0,0011 $ · 0,11 кредитов

Каждый запрос округляется до 0,01 кредита.

04

API

Вызовите DeepSeek V4 Flash с помощью вашего API-ключа Railwail. Используйте этот ID модели в запросе:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Установите ключ как RAILWAIL_API_KEYСоздать API-ключ
05

Спецификации

ID модели
deepseek-v4-flash
Разработчик
DeepSeek
Категория
Текст и чат
Входные данные
Текст
Выходные данные
Текст
Контекстное окно
1 048 575 токенов
Макс. выходные данные
384 000 токенов
Тарификация
По использованию (токены или время GPU)
Время выполнения (медиана)
1,3 s15 завершённых запросов на Railwail за последние 90 дней
Жизненный цикл
Снято с производства
Размер модели
284B total / 13B active per token
Лицензия
Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.
Запись в каталоге обновлена
23 сентября 2026 г.

Теги

  • deepseek
  • open-weights
  • moe
  • cost-efficient
  • long-context
  • 1m-context
06

Применение

Для чего это используется

  • Production RAG pipelines
  • High-throughput coding subagents
  • Bulk data extraction and classification
  • Cost-sensitive enterprise APIs
  • On-premise inference under tight cost budgets
  • Long-document summarisation at scale
  • Real-time chat backends
07

Часто задаваемые вопросы

Что такое DeepSeek V4 Flash?

DeepSeek V4 Flash — модель от DeepSeek в категории Текст и чат. На Railwail вы можете вызвать её через API Railwail с помощью ключа API.

Сколько стоит DeepSeek V4 Flash на Railwail?

На Railwail DeepSeek V4 Flash стоит 0,36 $ за 1M входных токенов и 1,44 $ за 1M выходных токенов. Вы платите за то, что фактически использует каждый запрос. Использование оплачивается предоплаченными кредитами; 1 кредит = 0,01 $.

Какой размер контекстного окна у DeepSeek V4 Flash?

Контекстное окно DeepSeek V4 Flash содержит 1 048 575 токенов. Ответ может быть до 384 000 токенов.

Насколько быстра DeepSeek V4 Flash?

На Railwail медианное время выполнения DeepSeek V4 Flash за последние 90 дней составило 1,3 s, на основе 15 завершённых запусков.

DeepSeek V4 Flash лучше, чем DeepSeek V4.1 Flash?

Это зависит от задачи. DeepSeek V4 Flash (DeepSeek) и DeepSeek V4.1 Flash (DeepSeek) — обе модели в категории Текст и чат. На странице сравнения показаны их цены и характеристики рядом.

Сравнить DeepSeek V4 Flash и DeepSeek V4.1 Flash

Как использовать DeepSeek V4 Flash через API?

Создайте ключ API Railwail и отправьте запрос с ID модели deepseek-v4-flash. Примеры кода для curl, Python и JavaScript находятся в разделе API на этой странице.

08

Сравнимые модели

Все в этой категории

Использовать DeepSeek V4 Flash через API

Один API-ключ для всех моделей на Railwail. Использование оплачивается из предоплаченных кредитов, 1 кредит = 0,01 $.