DeepSeek V4 Pro

Testo e chatNuovoDisponibile
di DeepSeekID modello: deepseek-v4-pro

DeepSeek's April 2026 flagship. 1.6T MoE / 49B active params, 1M context, rivals top closed-source models on STEM and coding at a fraction of the price.

Prezzo · 1M in/out
1,584 USD / 4,752 USD
Contesto
1.048.576 token
Max. output
384.000 token
Input → output
Testo → Testo
Sviluppatore
DeepSeek
Aggiornato
23 settembre 2026
01

Playground

Prova DeepSeek V4 Pro

Chat

1,584 USD/1M in
Prova DeepSeek V4 Pro

Invia un messaggio. La risposta arriva completa quando il modello ha finito (senza streaming).

Prompt di sistema
Lunghezza massima della risposta (token)

Questa esecuzione

al massimo 0,0049 USD · 0,49 crediti riservati

Fatturato in base ai token effettivamente utilizzati; la parte inutilizzata della prenotazione viene rimborsata.

Nuovo qui?

10 crediti gratuiti (0,10 USD) quando ti iscrivi con Google

Utilizzabile 24 ore dopo l'iscrizione, fino a 5 esecuzioni al giorno e al massimo 2 crediti per esecuzione. Altri metodi di accesso iniziano senza crediti. Sufficiente per 20 esecuzioni di questo modello.

02

Informazioni su DeepSeek V4 Pro

RiassuntoA partire da 23 settembre 2026

DeepSeek V4 Pro è un modello di DeepSeek nella categoria Testo e chat. Su Railwail, DeepSeek V4 Pro costa 1,584 USD per 1M token di input e 4,752 USD per 1M token di output. La finestra di contesto contiene 1.048.576 token e una risposta può essere lunga fino a 384.000 token.

Released April 24, 2026 as part of the DeepSeek V4 Preview, DeepSeek-V4-Pro is a 1.6T-parameter Mixture-of-Experts model with 49B activated parameters. Native 1M-token context, 384K max output. Tops open-weights leaderboards on Math/STEM/Coding (~81% SWE-bench Verified) and rivals frontier closed-source models. Best for: open-weights coding agents, long-document analysis, cost-efficient reasoning workloads.

Sfondo

Informazioni su DeepSeek AI

Fondato 2023 · Hangzhou, China

DeepSeek AI is a Chinese AI research lab founded in 2023 by Liang Wenfeng, the founder of the High-Flyer quantitative hedge fund. The lab is funded primarily by High-Flyer's profits and operates independently of major Chinese tech conglomerates. DeepSeek's mission is to build open frontier AI: every flagship model has been released with open weights and a permissive license. Major releases include DeepSeek LLM (late 2023), DeepSeek-V2 (May 2024, MoE), DeepSeek-V3 (December 2024, 671B MoE / 37B active), DeepSeek-R1 (January 2026 family, reasoning), DeepSeek V3.1 (early 2026) and the DeepSeek V4 family released as preview April 24, 2026. DeepSeek's models repeatedly top open-weights leaderboards on coding, math and reasoning at a fraction of the training cost claimed by Western labs, and the team is credited with popularising large-scale Reinforcement Learning from Verifiable Rewards.

Visita DeepSeek AI

Architettura

Sparse Mixture-of-Experts Transformer (frontier open-weights)

DeepSeek-V4-Pro was released April 24, 2026 as the flagship of the V4 family. It is a Sparse MoE Transformer with 1.6T total parameters and 49B activated per token, supporting a native 1M-token context window with up to 384K-token max output. The model was trained on the lab's expanded GPU cluster using DeepSeek's signature recipe: large-scale pretraining on a multi-trillion-token mixture of web text, code, books, scientific papers and curated math/STEM data, followed by extensive Reinforcement Learning from Verifiable Rewards (RLVR) on math, coding and tool-use trajectories. Architectural innovations introduced in V3 - Multi-head Latent Attention (MLA), DeepSeekMoE with fine-grained expert specialization and shared experts, and FP8 mixed-precision training - are retained and refined. V4 Pro is published with open weights under a permissive license and runs natively in both server and inference frameworks such as vLLM and SGLang. DeepSeek treats the V4 launch as a preview phase and has announced that the older deepseek-chat and deepseek-reasoner endpoints will be deprecated on July 24, 2026.

Parametri
1.6T total / 49B active per token
Contesto
1.048.576 token

Capacità

  • 1M token native context window with 384K max output
  • ~81% SWE-bench Verified - rivals top closed-source models
  • Top open-weights scores on Math/STEM/Coding benchmarks
  • 1.6T MoE / 49B active parameters
  • FP8-trained for compute efficiency
  • Multi-head Latent Attention for memory-efficient long context
  • Function calling and structured JSON output
  • Cache hit pricing at $0.0145 per million tokens enables cheap multi-turn agents
  • Available via DeepSeek API, OpenRouter, Together, Fireworks and self-hosted with vLLM/SGLang
  • Open weights under a permissive license
  • Best for: open-weights coding agents, long-document analysis, cost-efficient reasoning workloads, on-premise enterprise deployments.

Addestramento e licenza

Pretrained on a multi-trillion-token mixture of web text, code, books, scientific papers and curated math/STEM data. Post-training applies large-scale Reinforcement Learning from Verifiable Rewards (RLVR) on math, coding and tool-use tasks, plus supervised fine-tuning and instruction tuning. Knowledge cutoff approximately early 2026.

Licenza: Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.

Test di sicurezza: DeepSeek publishes model cards but provides limited external red-teaming. Safety filters are lighter than Western frontier labs; deployers are responsible for downstream alignment.

Limitazioni note

  • Light built-in safety alignment relative to Western frontier models
  • No native vision or audio input (text-only)
  • Pro variant requires substantial GPU resources to self-host
  • Older deepseek-chat / deepseek-reasoner endpoints will be deprecated July 24, 2026
  • Some Chinese-language safety constraints apply
03

Prezzi

Prezzi in dollari USA. L'utilizzo viene addebitato dai crediti prepagati.
Input1,584 USD / 1M token
Output4,752 USD / 1M token
  • Fatturato in base ai token che ogni richiesta utilizza effettivamente.
  • 1 credito = 0,01 USD

Calcolatore di costi

Calcolatore prezzi

/ richiesta
/ richiesta

Totale

0,40 USD

40 crediti

Per richiesta

0,004 USD · 0,4 crediti

Ogni richiesta viene arrotondata a 0,01 crediti.

04

API

Chiama DeepSeek V4 Pro con la tua chiave API Railwail. Usa questo ID modello nella richiesta:
curl https://railwail.com/api/v1/chat/completions \
  -H "Authorization: Bearer $RAILWAIL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-pro",
    "messages": [
      {
        "role": "user",
        "content": "Explain what a vector database is in two sentences."
      }
    ],
    "max_tokens": 1024
  }'
Imposta la tua chiave come RAILWAIL_API_KEYCrea chiave API
05

Specifiche

ID modello
deepseek-v4-pro
Sviluppatore
DeepSeek
Categoria
Testo e chat
Input
Testo
Output
Testo
Finestra di contesto
1.048.576 token
Output massimo
384.000 token
Fatturazione
In base all'utilizzo (token o tempo GPU)
Rilasciato
24 aprile 2026
Ciclo di vita
Versione attuale
Dimensione del modello
1.6T total / 49B active per token
Licenza
Open weights under a permissive license that allows commercial use. Hosted API access via deepseek.com.
Voce di catalogo aggiornata
23 settembre 2026

Parametri di input

Input e impostazioni dallo schema di input del modello. L'esempio nella sezione API mostra quali di essi l'API accetta.

  • promptObbligatorio

    User message

    Tipo: Testo
    Predefinito: –
    Valori consentiti: fino a 32.000 caratteri
  • top_p
    Tipo: Numero
    Predefinito: 1
    Valori consentiti: 0 a 1
  • stream
    Tipo: Sì/No
    Predefinito: false
    Valori consentiti: –
  • max_tokens
    Tipo: Numero intero
    Predefinito: 4096
    Valori consentiti: 1 a 32.000
  • temperature
    Tipo: Numero
    Predefinito: 0.7
    Valori consentiti: 0 a 2
  • system_prompt

    Optional system instruction

    Tipo: Testo
    Predefinito: –
    Valori consentiti: fino a 8000 caratteri

Etichette

  • deepseek
  • open-weights
  • moe
  • coding
  • reasoning
  • long-context
  • 1m-context
  • flagship
06

Casi d'uso

A cosa serve

  • Open-weights coding agents
  • Long-document analysis
  • On-premise enterprise deployments
  • Cost-efficient frontier reasoning workloads
  • Math and STEM tutoring backends
  • Research and academic use under permissive license
  • RAG over millions of tokens of context
07

Domande frequenti

Cos'è DeepSeek V4 Pro?

DeepSeek V4 Pro è un modello di DeepSeek nella categoria Testo e chat. Su Railwail puoi richiamarlo con una chiave API tramite l'API Railwail.

Quanto costa DeepSeek V4 Pro su Railwail?

Su Railwail, DeepSeek V4 Pro costa 1,584 USD per 1M token di input e 4,752 USD per 1M token di output. Ti viene addebitato ciò che ogni richiesta utilizza effettivamente. L'utilizzo viene pagato con crediti prepagati; 1 credito equivale a 0,01 USD.

Qual è la finestra di contesto di DeepSeek V4 Pro?

La finestra di contesto di DeepSeek V4 Pro contiene 1.048.576 token. Una risposta può essere lunga fino a 384.000 token.

Quanto è veloce DeepSeek V4 Pro?

Non ci sono ancora abbastanza esecuzioni misurate di DeepSeek V4 Pro su Railwail per indicare un tempo di esecuzione. Dipende dall'input, dalle impostazioni e dal carico presso il provider.

DeepSeek V4 Pro è migliore di Claude Fable 5.1?

Dipende dall'attività. DeepSeek V4 Pro (DeepSeek) e Claude Fable 5.1 (Anthropic) sono entrambi modelli nella categoria Testo e chat. La pagina di confronto mostra i loro prezzi e le specifiche affiancati.

Confronta DeepSeek V4 Pro e Claude Fable 5.1

Come uso DeepSeek V4 Pro tramite l'API?

Crea una chiave API Railwail e invia la tua richiesta con l'ID modello deepseek-v4-pro. Gli esempi di codice per curl, Python e JavaScript sono nella sezione API di questa pagina.

08

Modelli comparabili

Tutti in questa categoria

Usa DeepSeek V4 Pro tramite l'API

Una chiave API per tutti i modelli su Railwail. L'utilizzo viene addebitato da crediti prepagati, 1 credito = 0,01 USD.