DeepSeek V4 Flash

Texto y chatDescontinuadoDisponible
de DeepSeekID del modelo: deepseek-v4-flash

Efficiency-optimized variant of DeepSeek V4. 284B MoE / 13B active, 1M context, ultra-low pricing for high-throughput workloads.

Precio · 1M entrada/salida
USD 0.36 / USD 1.44
Contexto
1,048,575 tokens
Salida máxima
384,000 tokens
Entrada → Salida
Texto → Texto
Tiempo de ejecución (mediana)
1.3 s
Desarrollador
DeepSeek

El proveedor está descontinuando este modelo.

Versión más nueva disponible: DeepSeek V4.1 Flash

01

Playground

Probar DeepSeek V4 Flash

Chat

USD 0.36/1M entrada
Probar DeepSeek V4 Flash

Envía un mensaje. La respuesta llega completa cuando el modelo termina (sin transmisión en tiempo real).

Indicación del sistema
Longitud máxima de respuesta (tokens)

Esta ejecución

como máximo USD 0.0015 · 0.15 créditos reservados

Se facturan los tokens realmente utilizados; la parte no utilizada de la reserva se reembolsa.

¿Nuevo aquí?

10 créditos gratis (USD 0.10) cuando te registras con Google

Utilizable 24 horas después del registro, hasta 5 ejecuciones por día y como máximo 2 créditos por ejecución. Otros métodos de inicio de sesión comienzan sin créditos. Suficiente para 66 ejecuciones de este modelo.

02

Acerca de DeepSeek V4 Flash

ResumenA partir de 23 de septiembre de 2026

DeepSeek V4 Flash es un modelo de DeepSeek en la categoría Texto y chat. En Railwail, DeepSeek V4 Flash cuesta USD 0.36 por 1M tokens de entrada y USD 1.44 por 1M tokens de salida. La ventana de contexto contiene 1,048,575 tokens, y una respuesta puede tener hasta 384,000 tokens. Versión más reciente: DeepSeek V4.1 Flash.

DeepSeek-V4-Flash is the cost-efficient sibling of V4-Pro, released April 2026 as part of the V4 Preview. 284B total / 13B active MoE parameters with the same 1M-token context window. Designed for high-throughput agentic loops, RAG and batch tasks where latency and cost matter more than raw capability. Recommended for production agents, classification at scale, large-scale data extraction.

Fondo

Acerca de DeepSeek AI

Fundado en 2023 · Hangzhou, China

DeepSeek AI is a Chinese AI research lab founded in 2023 by Liang Wenfeng, founder of the High-Flyer quantitative hedge fund. The lab is funded primarily by High-Flyer's profits. Its mission is open frontier AI, with all flagship models released with open weights. Major releases include DeepSeek LLM (2023), DeepSeek-V2 (May 2024), DeepSeek-V3 (December 2024), DeepSeek-R1 (January 2026), DeepSeek V3.1 (early 2026) and the DeepSeek V4 family (April 24, 2026), comprising V4-Pro and V4-Flash. DeepSeek is credited with popularising large-scale Reinforcement Learning from Verifiable Rewards and consistently tops open-weights leaderboards.

Visitar DeepSeek AI

Arquitectura

Sparse Mixture-of-Experts Transformer (efficiency-optimized open-weights)

DeepSeek-V4-Flash was released April 24, 2026 as the efficiency-optimized sibling of V4-Pro. It is a Sparse MoE Transformer with 284B total parameters and 13B activated per token, retaining the full 1M-token native context window and 384K-token max output of the Pro variant at significantly lower inference cost. The model uses the same DeepSeek architectural stack: Multi-head Latent Attention (MLA), DeepSeekMoE with fine-grained expert specialization and shared experts, and FP8 mixed-precision training. Post-training combined supervised fine-tuning, RLVR on math/code/tool-use trajectories, and heavy distillation from the V4-Pro teacher model. V4 Flash is published with open weights under a permissive license and is designed for production-scale RAG, agentic loops and high-throughput workloads. At $0.112 input / $0.224 output per million tokens it undercuts every Western frontier model by an order of magnitude.

Parámetros
284B total / 13B active per token
Contexto
1,048,575 tokens

Capacidades

  • 1M token native context window with 384K max output
  • 284B MoE / 13B active parameters
  • Ultra-low pricing ($0.112 / $0.224 per million tokens)
  • Distilled from DeepSeek V4-Pro teacher model
  • FP8-trained for compute efficiency
  • Multi-head Latent Attention for memory-efficient long context
  • Function calling and structured JSON output
  • Strong on math, STEM and coding for its size
  • Available via DeepSeek API, OpenRouter, Together and self-hosted with vLLM/SGLang
  • Open weights under a permissive license
  • Best for: production agents, RAG pipelines, high-throughput data extraction, on-premise inference under tight cost budgets.

Entrenamiento y licencia

Pretrained on the same multi-trillion-token mixture as V4-Pro. Post-training combines supervised fine-tuning, RLVR and distillation from the V4-Pro teacher model. Knowledge cutoff approximately early 2026.

Licencia: Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.

Pruebas de seguridad: DeepSeek publishes model cards but provides limited external red-teaming. Safety filters are lighter than Western frontier labs; deployers are responsible for downstream alignment.

Limitaciones conocidas

  • Below V4-Pro on the hardest reasoning and coding benchmarks
  • Light built-in safety alignment relative to Western frontier models
  • No native vision or audio input (text-only)
  • Older deepseek-chat / deepseek-reasoner endpoints will be deprecated July 24, 2026
  • Some Chinese-language safety constraints apply
03

Precios

Precios en dólares estadounidenses. El uso se cobra con créditos prepagados.
EntradaUSD 0.36 / 1M tokens
SalidaUSD 1.44 / 1M tokens
  • Se facturan los tokens que cada solicitud realmente utiliza.
  • 1 crédito = USD 0.01

Calculadora de costos

Calculadora de precios

/ solicitud
/ solicitud

Total

USD 0.11

11 créditos

Por solicitud

USD 0.0011 · 0.11 créditos

Cada solicitud se redondea a 0.01 créditos.

04

API

Llama a DeepSeek V4 Flash con tu clave de API de Railwail. Usa este ID de modelo en la solicitud:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Establece tu clave como RAILWAIL_API_KEYCrear clave API
05

Especificaciones

ID del modelo
deepseek-v4-flash
Desarrollador
DeepSeek
Categoría
Texto y chat
Entrada
Texto
Salida
Texto
Ventana de contexto
1,048,575 tokens
Salida máxima
384,000 tokens
Facturación
Por uso (tokens o tiempo de GPU)
Tiempo de ejecución (mediana)
1.3 s15 ejecuciones completadas en Railwail en los últimos 90 días
Ciclo de vida
Descontinuado
Tamaño del modelo
284B total / 13B active per token
Licencia
Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.
Entrada del catálogo actualizada
23 de septiembre de 2026

Etiquetas

  • deepseek
  • open-weights
  • moe
  • cost-efficient
  • long-context
  • 1m-context
06

Casos de uso

Para qué se utiliza

  • Production RAG pipelines
  • High-throughput coding subagents
  • Bulk data extraction and classification
  • Cost-sensitive enterprise APIs
  • On-premise inference under tight cost budgets
  • Long-document summarisation at scale
  • Real-time chat backends
07

Preguntas frecuentes

¿Qué es DeepSeek V4 Flash?

DeepSeek V4 Flash es un modelo de DeepSeek en la categoría Texto y chat. En Railwail puedes llamarlo con una clave API a través de la API de Railwail.

¿Cuánto cuesta DeepSeek V4 Flash en Railwail?

En Railwail, DeepSeek V4 Flash cuesta USD 0.36 por 1M tokens de entrada y USD 1.44 por 1M tokens de salida. Se te cobra por lo que cada solicitud realmente usa. El uso se paga con créditos prepagados; 1 crédito equivale a USD 0.01.

¿Cuál es la ventana de contexto de DeepSeek V4 Flash?

La ventana de contexto de DeepSeek V4 Flash contiene 1,048,575 tokens. Una respuesta puede tener hasta 384,000 tokens.

¿Qué tan rápido es DeepSeek V4 Flash?

En Railwail, el tiempo de ejecución mediano de DeepSeek V4 Flash en los últimos 90 días fue 1.3 s, basado en 15 ejecuciones completadas.

¿Es DeepSeek V4 Flash mejor que DeepSeek V4.1 Flash?

Eso depende de la tarea. DeepSeek V4 Flash (DeepSeek) y DeepSeek V4.1 Flash (DeepSeek) son ambos modelos en la categoría Texto y chat. La página de comparación muestra sus precios y especificaciones lado a lado.

Comparar DeepSeek V4 Flash y DeepSeek V4.1 Flash

¿Cómo uso DeepSeek V4 Flash a través de la API?

Crea una clave API de Railwail y envía tu solicitud con el ID de modelo deepseek-v4-flash. Los ejemplos de código para curl, Python y JavaScript están en la sección API de esta página.

08

Modelos comparables

Todos en esta categoría

Usar DeepSeek V4 Flash a través de la API

Una clave API para todos los modelos en Railwail. El uso se cobra con créditos prepagados, 1 crédito = USD 0.01.