DeepSeek V3.1

Szöveg és csevegésKivezetveNem elérhető
DeepSeek általModell-azonosító: deepseek-v3-1

DeepSeek's refreshed V3.1 release. 671B MoE / 37B active. Tops open-weights leaderboards on coding and reasoning.

Állapot
Nem elérhető
Kontextus
131 072 token
Max. kimenet
8192 token
Bemenet → kimenet
Szöveg → Szöveg
Fejlesztő
DeepSeek
Frissítve
2026. szeptember 23.

DeepSeek V3.1 jelenleg nem elérhető

Az oldal részleteit továbbra is elolvashatja. Az alábbi elérhető alternatívák közül válasszon egy hasonló modell azonnali futtatásához.

Ugrás az alternatívákhoz

A szolgáltató kivezette ezt a modellt.

Újabb verzió érhető el: DeepSeek V4.1 Flash

01
02

Playground

Próbáld ki: DeepSeek V3.1

Csevegés

Jelenleg nem elérhető

Jelenleg nem elérhető.

A játszótér letiltva van. Hasonló modelleket találsz ugyanabban a kategóriában: Alternatívák megtekintése

Próbáld ki: DeepSeek V3.1

Írjon üzenetet. A válasz teljes egészében megérkezik, amikor a modell kész (nincs streaming).

Rendszer-prompt
Max. válaszlength (tokenek)

Ez a futtatás

Nincs ár – jelenleg nem elérhető.

Új vagy itt?

10 ingyenes kredit (0,10 USD) Google-fiókkal való regisztrációkor

Használható 24 órával a regisztráció után, naponta legfeljebb 5 futtatás és futtatásonként legfeljebb 2 kredit. Más bejelentkezési módok kreditek nélkül indulnak.

03

A DeepSeek V3.1 névjegye

Röviden2026. szeptember 23. szerint

A DeepSeek V3.1 a DeepSeek által fejlesztett modell a Szöveg és csevegés kategóriában. A DeepSeek V3.1 jelenleg nem érhető el a Railwail-en. A kontextablak 131 072 tokent tartalmaz, egy válasz pedig legfeljebb 8192 token hosszú lehet. Újabb verzió: DeepSeek V4.1 Flash.

Háttér

A DeepSeek névjegye

Alapítva: 2023 · Hangzhou, China

DeepSeek AI was founded in July 2023 in Hangzhou by Liang Wenfeng, also co-founder of the High-Flyer quantitative hedge fund. The fund's pre-export-control GPU cluster financed DeepSeek's training runs. The lab is known for transparent technical reports and an aggressive open-weights strategy under MIT license. Releases include DeepSeek Coder (Nov 2023), DeepSeek LLM 67B (Jan 2024), DeepSeekMath with GRPO (Feb 2024), DeepSeek V2 introducing Multi-head Latent Attention (May 2024), DeepSeek V3 in December 2024 trained for ~$5.6M of GPU-hours, DeepSeek R1 in January 2025 and DeepSeek V3.1 in 2025 as an incremental update consolidating the base model and the R1 reasoning capabilities into a unified hybrid model. The company has roughly 200 researchers and is privately backed by High-Flyer rather than venture capital. Its V3/R1 release triggered a global re-evaluation of frontier-AI training economics and a notable stock-market move in late January 2025.

DeepSeek meglátogatása

Architektúra

Sparse Mixture-of-Experts Transformer (hybrid base + thinking modes)

DeepSeek V3.1 is a 2025 update of the V3 base that unifies chat (non-thinking) and reasoning (thinking) modes into a single hybrid checkpoint. It retains the V3 architecture - a Sparse MoE Transformer with 671B total and 37B active parameters using DeepSeekMoE routing and Multi-head Latent Attention - but expands the pretraining corpus and updates the post-training recipe. According to DeepSeek's release notes, V3.1 was continually pretrained on ~840B additional tokens of long-context data, extending effective context handling and improving long-document recall within the 128K window. Post-training merged the V3 chat data with R1-style long-CoT reasoning data plus tool-use and agentic trajectories. V3.1 exposes two operating modes selected via the chat template: 'non-thinking' (V3-style fast responses) and 'thinking' (R1-style chain-of-thought before the answer), letting developers choose per request. Tool use and function calling are first-class and improved over both V3 and R1. The model also includes targeted strengthening on coding, agent benchmarks (SWE-bench, Terminal-Bench), and search-augmented reasoning. Weights are released under MIT license and the official DeepSeek API hosts both V3.1 and V3.1-Terminus checkpoints.

Paraméterek
671B total, 37B active per token (extended for V3.1)
Kontextus
128 000 token

Képességek

  • Hybrid model: switchable thinking / non-thinking modes in one checkpoint
  • 671B-parameter MoE with 37B active per token
  • 128K context window, retrained on ~840B additional long-context tokens
  • Strong agentic and tool-use performance on SWE-bench Verified and Terminal-Bench
  • Function calling and parallel tool calls
  • Long-CoT reasoning inherited from R1
  • Open weights under MIT license
  • DeepSeek API approximately 1/20th the cost of GPT-4o-class models
  • Compatible with vLLM, SGLang, llama.cpp, HuggingFace
  • Improved code editing and diff-format generation
  • Best for: budget-conscious agentic workloads, coding, hybrid reasoning, on-prem enterprise.

Képzés és licenc

Built on V3's 14.8T-token base, then continually pretrained on roughly 840B additional tokens biased toward long-context documents and code. Post-training combines V3 chat data with R1-style long-CoT and agentic tool-use trajectories.

Licenc: MIT license for weights, code and tokenizer; commercial use permitted.

Biztonsági tesztelés: Limited published safety evaluations. As with V3 and R1, politically sensitive topics aligned to Chinese regulations are filtered while general-purpose refusal rates remain low.

Ismert korlátozások

  • Sensitive Chinese political topics filtered
  • Large memory footprint requires multi-GPU inference
  • Text-only inputs (no native vision)
  • Knowledge cutoff approximately late 2024
  • Hybrid mode switching adds prompt-template complexity
04

Árak

Jelenleg nem elérhető. Jelenleg nincs ár ehhez a modellhez, ezért nem futtatható.

05

API

Hívja meg a DeepSeek V3.1 modellt a Railwail API-kulcsával. Használja ezt a modell-azonosítót a kérésben:

Jelenleg nem elérhető

A modellnek nincs ellenőrzött ára vagy deaktiválva van; az API-hívások visszautasítottak.

06

Specifikációk

Modell-azonosító
deepseek-v3-1
Fejlesztő
DeepSeek
Bemenet
Szöveg
Kimenet
Szöveg
Kontextusablak
131 072 token
Max. kimenet
8192 token
Életciklus
Kivezetve
Modell mérete
671B total, 37B active per token (extended for V3.1)
Licenc
MIT license for weights, code and tokenizer; commercial use permitted.
Katalógus bejegyzés frissítve
2026. szeptember 23.

Címkék

  • deepseek
  • open-weights
  • moe
  • coding
  • reasoning
07

Felhasználási esetek

Mire használják

  • Hybrid agentic and chat workloads
  • Coding agents with tool use
  • Cost-sensitive enterprise deployments
  • Search-augmented reasoning
  • Long-document analysis
  • On-prem multilingual chat
08

Gyakran ismételt kérdések

Mi az a DeepSeek V3.1?

A DeepSeek V3.1 a DeepSeek által fejlesztett modell a Szöveg és csevegés kategóriában. A Railwail katalógusában szerepel, de jelenleg nem futtatható.

Mennyibe kerül a DeepSeek V3.1 a Railwail-on?

A DeepSeek V3.1 jelenleg nem futtatható a Railwail-on, ezért nincs aktuális ár. Az elérhető alternatívák árakkal az oldal alább találhatók.

Mekkora a DeepSeek V3.1 kontextablaka?

A DeepSeek V3.1 kontextablaka 131 072 tokent tartalmaz. Egy válasz legfeljebb 8192 token hosszú lehet.

Milyen gyors a DeepSeek V3.1?

A DeepSeek V3.1-nek még nincs elég mért futtatása a Railwail-on ahhoz, hogy futási időt adjunk meg. Ez a bemenettől, a beállításoktól és a szolgáltató terhelésétől függ.

A DeepSeek V3.1 jobb, mint a DeepSeek V4.1 Flash?

Ez a feladattól függ. A DeepSeek V3.1 (DeepSeek) és a DeepSeek V4.1 Flash (DeepSeek) egyaránt modellek a Szöveg és csevegés kategóriában. Az összehasonlítás oldal az árakat és specifikációkat egymás mellett mutatja.

DeepSeek V3.1 és DeepSeek V4.1 Flash összehasonlítása

Használhatom a DeepSeek V3.1-t most?

Jelenleg nem elérhető. Az oldal online marad; az ugyanabból a kategóriából elérhető alternatívák az oldal alább találhatók.

Összes modell egy API-n keresztül

Egy API-kulcs a Railwail összes modelljéhez. A használat előre feltöltött kreditek alapján kerül felszámításra, 1 kredit = 0,01 USD.