Qwen 3 235B Instruct

Text a chatNedostupné
od Alibaba / QwenID modelu: qwen-3-235b

Alibaba's Qwen 3 flagship MoE: 235B total / 22B active. Strong reasoning and tool use, open-weights.

Stav
Nedostupné
Kontext
131 072 tokenov
Max. výstup
16 384 tokenov
Vstup → výstup
Text → Text
Vývojár
Alibaba / Qwen
Aktualizované
25. júna 2026

Qwen 3 235B Instruct nie je momentálne dostupný

Momentálne nedostupné: tento model bol deaktivovaný.

Podrobnosti na tejto stránke si môžete prečítať. Vyberte si jednu z dostupných alternatív nižšie a spustite porovnateľný model hneď.

Prejsť na alternatívy
01

Porovnateľné modely

Všetky v tejto kategórii
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    12,00 USD/1M vstup

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    6,00 USD/1M vstup

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    4,80 USD/1M vstup

02

Playground

Vyskúšajte Qwen 3 235B Instruct

Chat

Momentálne nedostupné

Momentálne nedostupné: tento model bol deaktivovaný.

Playground je vypnutý. Porovnateľné modely nájdete v tej istej kategórii: Pozrieť alternatívy

Vyskúšajte Qwen 3 235B Instruct

Pošli správu. Odpoveď príde úplne, keď je model hotový (bez streamovania).

Systémový prompt
Maximálna dĺžka odpovede (tokeny)

Tento beh

Bez ceny – momentálne nedostupné.

Nový tu?

5 bezplatných credits (0,05 USD) pri registrácii cez Google

Použiteľné 24 hodín po registrácii, až 5 behov za deň a maximálne 2 credits za beh. Ostatné spôsoby prihlásenia sa spúšťajú bez credits.

03

O Qwen 3 235B Instruct

StručneK 25. júna 2026

Qwen 3 235B Instruct je model od Alibaba / Qwen v kategórii Text a chat. Qwen 3 235B Instruct nie je v súčasnosti dostupný na Railwail. Kontextné okno obsahuje 131 072 tokenov a jedna odpoveď môže byť dlhá až 16 384 tokenov.

Pozadie

O Alibaba Cloud (Qwen team)

Založené 2009 · Hangzhou, China

The Qwen team inside Alibaba Cloud has shipped one of the most prolific open-weight model lines in industry, starting with Qwen-7B (Aug 2023, first Chinese open-weight foundation model from a hyperscaler), through Qwen 1.5 (Feb 2024), Qwen2 (Jun 2024), Qwen2.5 (Sep 2024, sizes 0.5B-72B plus Coder/Math/VL), Qwen2.5-Max closed-weight flagship (Jan 2025) and the Qwen3 family in April-May 2025. Qwen3 introduced a uniform 'hybrid thinking' approach across the family, letting developers toggle between fast direct answers and long chain-of-thought reasoning in the same checkpoint. The 235B-A22B MoE flagship anchors the family alongside dense variants from 0.6B to 32B. The team is led by Junyang Lin and has published over a dozen technical reports. Models ship under the Apache 2.0 license starting with Qwen3, a substantial liberalisation over the earlier Tongyi Qianwen LICENSE. Alibaba Cloud, founded in 2009, is the largest cloud provider in China and hosts the Qwen models on its Model Studio service while also distributing weights freely on HuggingFace, ModelScope and GitHub.

Navštíviť Alibaba Cloud (Qwen team)

Architektúra

Sparse Mixture-of-Experts Transformer (hybrid thinking)

Qwen3-235B-A22B is the flagship Mixture-of-Experts model of the Qwen3 family, released by Alibaba's Qwen team on 29 April 2025 with weights under Apache 2.0. The architecture is a Sparse MoE Transformer with 235 billion total parameters and 22 billion active per token (128 experts, 8 selected per token), 94 layers, and Grouped Query Attention (64 query heads / 4 KV heads). It supports a 128K native context window extended to 256K via YaRN scaling. The model was pretrained on approximately 36 trillion tokens spanning 119 languages with strong Chinese, English and code coverage, more than doubling the 18T-token corpus used for Qwen2.5. Pretraining was performed in three stages with progressively longer context lengths and improved data filtering. Post-training applied a four-stage pipeline: long-CoT cold start, RL on reasoning tasks, integration of thinking and non-thinking modes via mixed SFT, and a final general-purpose RL stage. The resulting model exposes hybrid thinking, toggled via the chat template, where the same checkpoint can produce either an R1-style chain-of-thought before the answer or a fast direct response. Qwen3-235B leads several open-weight benchmarks including AIME 2025, LiveCodeBench, ArenaHard and BFCL agent-eval as of release.

Parametre
235B total, 22B active per token
Kontext
262 144 tokenov

Schopnosti

  • 235B-parameter MoE with 22B active per token (128 experts, 8 selected)
  • Hybrid thinking mode toggle (CoT or direct answer in one checkpoint)
  • Pretrained on ~36T tokens across 119 languages
  • 256K context window with YaRN scaling
  • Apache 2.0 license, fully commercial
  • Top open-weight scores on AIME 2025, LiveCodeBench, ArenaHard, BFCL
  • Function calling, MCP server support, parallel tool calls
  • Specialised siblings: Qwen3-Coder, Qwen3-Math, Qwen3-VL
  • Compatible with vLLM, SGLang, llama.cpp, Ollama, MLX, HuggingFace
  • Broad multilingual coverage with strong Chinese/English/Japanese/Korean performance
  • Best for: open-weight reasoning, agentic workloads, multilingual chat, on-prem enterprise.

Tréning a licencia

Pretrained on approximately 36 trillion tokens covering 119 languages with strong Chinese and English emphasis, code repositories and scientific content. Post-training uses a four-stage pipeline: long-CoT cold start, reasoning RL, mixed SFT integrating thinking/non-thinking modes, and a final general-purpose RL stage.

Licencia: Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).

Bezpečnostné testy: Standard SFT+DPO safety alignment plus general RL safety stage. Filters Chinese politically sensitive topics. Limited third-party safety evaluations published.

Známe obmedzenia

  • Filters Chinese political topics
  • Large memory footprint requires multi-GPU inference for FP16
  • Vision requires separate Qwen3-VL checkpoint
  • Knowledge cutoff approximately early 2025
  • Long context >128K degrades on some recall tasks
04

Ceny

Momentálne nedostupné: tento model bol deaktivovaný. V súčasnosti nie je cena za tento model, preto ho nie je možné spustiť.

05

API

Zavolajte Qwen 3 235B Instruct s vaším API kľúčom Railwail. V požiadavke použite toto ID modelu:

Momentálne nedostupné

Model nemá overenú cenu alebo je deaktivovaný; volania API sú odmietnuté.

06

Špecifikácie

ID modelu
qwen-3-235b
Vývojár
Alibaba / Qwen
Kategória
Text a chat
Vstup
Text
Výstup
Text
Kontextové okno
131 072 tokenov
Max. výstup
16 384 tokenov
Veľkosť modelu
235B total, 22B active per token
Licencia
Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).
Katalógová položka aktualizovaná
25. júna 2026

Vstupné parametre

Vstupy a nastavenia zo vstupnej schémy modelu. Príklad v sekcii API ukazuje, ktoré z nich API akceptuje.

  • promptpovinné

    User message

    Typ: Text
    Predvolené:
    Povolené hodnoty: až 16 000 znakov
  • top_p
    Typ: Číslo
    Predvolené: 1
    Povolené hodnoty: 0 až 1
  • stream
    Typ: Áno/Nie
    Predvolené: false
    Povolené hodnoty:
  • max_tokens
    Typ: Celé číslo
    Predvolené: 2048
    Povolené hodnoty: 1 až 16 384
  • temperature
    Typ: Číslo
    Predvolené: 0.7
    Povolené hodnoty: 0 až 2
  • system_prompt

    Optional system instruction

    Typ: Text
    Predvolené:
    Povolené hodnoty: až 8 000 znakov

Značky

  • qwen
  • alibaba
  • moe
  • open-weights
  • flagship
07

Prípady použitia

Na čo sa používa

  • Open-weight reasoning workloads
  • Agentic tool-using applications
  • Multilingual chat across 100+ languages
  • On-prem enterprise deployments
  • Fine-tuning base for vertical models
  • Apache-2.0-required commercial products
08

Často kladené otázky

Čo je Qwen 3 235B Instruct?

Qwen 3 235B Instruct je model od Alibaba / Qwen v kategórii Text a chat. Je uvedený na Railwail, ale momentálne ho nie je možné spustiť.

Koľko stojí Qwen 3 235B Instruct na Railwail?

Qwen 3 235B Instruct nie je momentálne možné spustiť na Railwail, preto nie je aktuálna cena. Dostupné alternatívy s cenami sú uvedené nižšie na tejto stránke.

Aké je kontextné okno Qwen 3 235B Instruct?

Kontextné okno Qwen 3 235B Instruct obsahuje 131 072 tokenov. Jedna odpoveď môže byť dlhá až 16 384 tokenov.

Ako rýchly je Qwen 3 235B Instruct?

Pre Qwen 3 235B Instruct je na Railwail zatiaľ príliš málo meraných spustení na určenie doby spustenia. Závisí to od vstupu, nastavení a zaťaženia u poskytovateľa.

Je Qwen 3 235B Instruct lepší ako Claude Fable 5.1?

Závisí to od úlohy. Qwen 3 235B Instruct (Alibaba / Qwen) a Claude Fable 5.1 (Anthropic) sú oba modely v kategórii Text a chat. Stránka porovnania zobrazuje ich ceny a špecifikácie vedľa seba.

Porovnať Qwen 3 235B Instruct a Claude Fable 5.1

Môžem Qwen 3 235B Instruct používať práve teraz?

Momentálne nedostupné: tento model bol deaktivovaný. Stránka zostáva online; dostupné alternatívy z tej istej kategórie sú uvedené nižšie.

Všetky modely cez jedno API

Jeden API kľúč pre všetky modely na Railwail. Použitie sa účtuje z predplateného kreditu, 1 kredit = 0,01 USD.