DeepSeek V3.1

Tekst & chatStopgezetNiet beschikbaar
van DeepSeekModel-ID: deepseek-v3-1

DeepSeek's refreshed V3.1 release. 671B MoE / 37B active. Tops open-weights leaderboards on coding and reasoning.

Status
Niet beschikbaar
Context
131.072 tokens
Max. uitvoer
8.192 tokens
Invoer → Uitvoer
Tekst → Tekst
Ontwikkelaar
DeepSeek
Bijgewerkt
23 september 2026

DeepSeek V3.1 is momenteel niet beschikbaar

Je kunt de details op deze pagina nog steeds lezen. Kies een van de beschikbare alternatieven hieronder om direct een vergelijkbaar model uit te voeren.

Naar alternatieven

De provider heeft dit model stopgezet.

Nieuwere versie beschikbaar: DeepSeek V4.1 Flash

01

Vergelijkbare modellen

Alle in deze categorie
02

Playground

DeepSeek V3.1 proberen

Chat

Momenteel niet beschikbaar

Momenteel niet beschikbaar.

De playground is uitgeschakeld. Vergelijkbare modellen vind je in dezelfde categorie: Alternatieven bekijken

DeepSeek V3.1 proberen

Stuur een bericht. Het antwoord komt volledig binnen zodra het model klaar is (geen streaming).

Systeemprompt
Max. antwoordlengte (tokens)

Deze uitvoering

Geen prijs – momenteel niet beschikbaar.

Nieuw hier?

10 gratis credits (US$ 0,10) wanneer je je aanmeldt met Google

Bruikbaar 24 uur na aanmelding, tot 5 uitvoeringen per dag en maximaal 2 credits per uitvoering. Andere aanmeldmethoden starten zonder credits.

03

Over DeepSeek V3.1

SamengevatPer 23 september 2026

DeepSeek V3.1 is een model van DeepSeek in de categorie Tekst & chat. DeepSeek V3.1 is momenteel niet beschikbaar op Railwail. Het contextvenster bevat 131.072 tokens, en een antwoord kan tot 8.192 tokens lang zijn. Nieuwere versie: DeepSeek V4.1 Flash.

Achtergrond

Over DeepSeek

Opgericht 2023 · Hangzhou, China

DeepSeek AI was founded in July 2023 in Hangzhou by Liang Wenfeng, also co-founder of the High-Flyer quantitative hedge fund. The fund's pre-export-control GPU cluster financed DeepSeek's training runs. The lab is known for transparent technical reports and an aggressive open-weights strategy under MIT license. Releases include DeepSeek Coder (Nov 2023), DeepSeek LLM 67B (Jan 2024), DeepSeekMath with GRPO (Feb 2024), DeepSeek V2 introducing Multi-head Latent Attention (May 2024), DeepSeek V3 in December 2024 trained for ~$5.6M of GPU-hours, DeepSeek R1 in January 2025 and DeepSeek V3.1 in 2025 as an incremental update consolidating the base model and the R1 reasoning capabilities into a unified hybrid model. The company has roughly 200 researchers and is privately backed by High-Flyer rather than venture capital. Its V3/R1 release triggered a global re-evaluation of frontier-AI training economics and a notable stock-market move in late January 2025.

DeepSeek bezoeken

Architectuur

Sparse Mixture-of-Experts Transformer (hybrid base + thinking modes)

DeepSeek V3.1 is a 2025 update of the V3 base that unifies chat (non-thinking) and reasoning (thinking) modes into a single hybrid checkpoint. It retains the V3 architecture - a Sparse MoE Transformer with 671B total and 37B active parameters using DeepSeekMoE routing and Multi-head Latent Attention - but expands the pretraining corpus and updates the post-training recipe. According to DeepSeek's release notes, V3.1 was continually pretrained on ~840B additional tokens of long-context data, extending effective context handling and improving long-document recall within the 128K window. Post-training merged the V3 chat data with R1-style long-CoT reasoning data plus tool-use and agentic trajectories. V3.1 exposes two operating modes selected via the chat template: 'non-thinking' (V3-style fast responses) and 'thinking' (R1-style chain-of-thought before the answer), letting developers choose per request. Tool use and function calling are first-class and improved over both V3 and R1. The model also includes targeted strengthening on coding, agent benchmarks (SWE-bench, Terminal-Bench), and search-augmented reasoning. Weights are released under MIT license and the official DeepSeek API hosts both V3.1 and V3.1-Terminus checkpoints.

Parameters
671B total, 37B active per token (extended for V3.1)
Context
128.000 tokens

Mogelijkheden

  • Hybrid model: switchable thinking / non-thinking modes in one checkpoint
  • 671B-parameter MoE with 37B active per token
  • 128K context window, retrained on ~840B additional long-context tokens
  • Strong agentic and tool-use performance on SWE-bench Verified and Terminal-Bench
  • Function calling and parallel tool calls
  • Long-CoT reasoning inherited from R1
  • Open weights under MIT license
  • DeepSeek API approximately 1/20th the cost of GPT-4o-class models
  • Compatible with vLLM, SGLang, llama.cpp, HuggingFace
  • Improved code editing and diff-format generation
  • Best for: budget-conscious agentic workloads, coding, hybrid reasoning, on-prem enterprise.

Training & licentie

Built on V3's 14.8T-token base, then continually pretrained on roughly 840B additional tokens biased toward long-context documents and code. Post-training combines V3 chat data with R1-style long-CoT and agentic tool-use trajectories.

Licentie: MIT license for weights, code and tokenizer; commercial use permitted.

Veiligheidstests: Limited published safety evaluations. As with V3 and R1, politically sensitive topics aligned to Chinese regulations are filtered while general-purpose refusal rates remain low.

Bekende beperkingen

  • Sensitive Chinese political topics filtered
  • Large memory footprint requires multi-GPU inference
  • Text-only inputs (no native vision)
  • Knowledge cutoff approximately late 2024
  • Hybrid mode switching adds prompt-template complexity
04

Prijzen

Momenteel niet beschikbaar. Er is momenteel geen prijs voor dit model, dus het kan niet worden uitgevoerd.

05

API

Roep DeepSeek V3.1 aan met je Railwail API-sleutel. Gebruik deze model-ID in het verzoek:

Momenteel niet beschikbaar

Het model heeft geen geverifieerde prijs of is gedeactiveerd; API-aanroepen worden geweigerd.

06

Specificaties

Model-ID
deepseek-v3-1
Ontwikkelaar
DeepSeek
Categorie
Tekst & chat
Invoer
Tekst
Uitvoer
Tekst
Contextvenster
131.072 tokens
Max. uitvoer
8.192 tokens
Levenscyclus
Stopgezet
Modelgrootte
671B total, 37B active per token (extended for V3.1)
Licentie
MIT license for weights, code and tokenizer; commercial use permitted.
Catalogusitem bijgewerkt
23 september 2026

Tags

  • deepseek
  • open-weights
  • moe
  • coding
  • reasoning
07

Gebruiksscenario's

Waarvoor het wordt gebruikt

  • Hybrid agentic and chat workloads
  • Coding agents with tool use
  • Cost-sensitive enterprise deployments
  • Search-augmented reasoning
  • Long-document analysis
  • On-prem multilingual chat
08

Veelgestelde vragen

Wat is DeepSeek V3.1?

DeepSeek V3.1 is een model van DeepSeek in de categorie Tekst & chat. Het staat in de Railwail-catalogus, maar kan momenteel niet worden uitgevoerd.

Hoeveel kost DeepSeek V3.1 op Railwail?

DeepSeek V3.1 kan momenteel niet op Railwail worden uitgevoerd, dus er is geen huidige prijs. Beschikbare alternatieven met prijzen staan verderop op deze pagina.

Wat is het contextvenster van DeepSeek V3.1?

Het contextvenster van DeepSeek V3.1 bevat 131.072 tokens. Een antwoord kan tot 8.192 tokens lang zijn.

Hoe snel is DeepSeek V3.1?

Er zijn nog niet genoeg gemeten runs van DeepSeek V3.1 op Railwail om een uitvoeringstijd op te geven. Dit hangt af van de invoer, de instellingen en de belasting bij de provider.

Is DeepSeek V3.1 beter dan DeepSeek V4.1 Flash?

Dat hangt van de taak af. DeepSeek V3.1 (DeepSeek) en DeepSeek V4.1 Flash (DeepSeek) zijn beide modellen in de categorie Tekst & chat. De vergelijkingspagina toont hun prijzen en specificaties naast elkaar.

DeepSeek V3.1 en DeepSeek V4.1 Flash vergelijken

Kan ik DeepSeek V3.1 nu gebruiken?

Momenteel niet beschikbaar. De pagina blijft online; beschikbare alternatieven uit dezelfde categorie staan verderop.

Alle modellen via één API

Één API-sleutel voor elk model op Railwail. Gebruik wordt afgerekend via vooraf gekochte credits, 1 credit = US$ 0,01.