Qwen 3 235B Instruct

Tekst & chatNiet beschikbaar
van Alibaba / QwenModel-ID: qwen-3-235b

Alibaba's Qwen 3 flagship MoE: 235B total / 22B active. Strong reasoning and tool use, open-weights.

Status
Niet beschikbaar
Context
131.072 tokens
Max. uitvoer
16.384 tokens
Invoer → Uitvoer
Tekst → Tekst
Ontwikkelaar
Alibaba / Qwen
Bijgewerkt
25 juni 2026

Qwen 3 235B Instruct is momenteel niet beschikbaar

Momenteel niet beschikbaar: dit model is gedeactiveerd.

Je kunt de details op deze pagina nog steeds lezen. Kies een van de beschikbare alternatieven hieronder om direct een vergelijkbaar model uit te voeren.

Naar alternatieven
01

Vergelijkbare modellen

Alle in deze categorie
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    US$ 12,00/1M in

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    US$ 6,00/1M in

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    US$ 4,80/1M in

02

Playground

Qwen 3 235B Instruct proberen

Chat

Momenteel niet beschikbaar

Momenteel niet beschikbaar: dit model is gedeactiveerd.

De playground is uitgeschakeld. Vergelijkbare modellen vind je in dezelfde categorie: Alternatieven bekijken

Qwen 3 235B Instruct proberen

Stuur een bericht. Het antwoord komt volledig binnen zodra het model klaar is (geen streaming).

Systeemprompt
Max. antwoordlengte (tokens)

Deze uitvoering

Geen prijs – momenteel niet beschikbaar.

Nieuw hier?

5 gratis credits (US$ 0,05) wanneer je je aanmeldt met Google

Bruikbaar 24 uur na aanmelding, tot 5 uitvoeringen per dag en maximaal 2 credits per uitvoering. Andere aanmeldmethoden starten zonder credits.

03

Over Qwen 3 235B Instruct

SamengevatPer 25 juni 2026

Qwen 3 235B Instruct is een model van Alibaba / Qwen in de categorie Tekst & chat. Qwen 3 235B Instruct is momenteel niet beschikbaar op Railwail. Het contextvenster bevat 131.072 tokens, en een antwoord kan tot 16.384 tokens lang zijn.

Achtergrond

Over Alibaba Cloud (Qwen team)

Opgericht 2009 · Hangzhou, China

The Qwen team inside Alibaba Cloud has shipped one of the most prolific open-weight model lines in industry, starting with Qwen-7B (Aug 2023, first Chinese open-weight foundation model from a hyperscaler), through Qwen 1.5 (Feb 2024), Qwen2 (Jun 2024), Qwen2.5 (Sep 2024, sizes 0.5B-72B plus Coder/Math/VL), Qwen2.5-Max closed-weight flagship (Jan 2025) and the Qwen3 family in April-May 2025. Qwen3 introduced a uniform 'hybrid thinking' approach across the family, letting developers toggle between fast direct answers and long chain-of-thought reasoning in the same checkpoint. The 235B-A22B MoE flagship anchors the family alongside dense variants from 0.6B to 32B. The team is led by Junyang Lin and has published over a dozen technical reports. Models ship under the Apache 2.0 license starting with Qwen3, a substantial liberalisation over the earlier Tongyi Qianwen LICENSE. Alibaba Cloud, founded in 2009, is the largest cloud provider in China and hosts the Qwen models on its Model Studio service while also distributing weights freely on HuggingFace, ModelScope and GitHub.

Alibaba Cloud (Qwen team) bezoeken

Architectuur

Sparse Mixture-of-Experts Transformer (hybrid thinking)

Qwen3-235B-A22B is the flagship Mixture-of-Experts model of the Qwen3 family, released by Alibaba's Qwen team on 29 April 2025 with weights under Apache 2.0. The architecture is a Sparse MoE Transformer with 235 billion total parameters and 22 billion active per token (128 experts, 8 selected per token), 94 layers, and Grouped Query Attention (64 query heads / 4 KV heads). It supports a 128K native context window extended to 256K via YaRN scaling. The model was pretrained on approximately 36 trillion tokens spanning 119 languages with strong Chinese, English and code coverage, more than doubling the 18T-token corpus used for Qwen2.5. Pretraining was performed in three stages with progressively longer context lengths and improved data filtering. Post-training applied a four-stage pipeline: long-CoT cold start, RL on reasoning tasks, integration of thinking and non-thinking modes via mixed SFT, and a final general-purpose RL stage. The resulting model exposes hybrid thinking, toggled via the chat template, where the same checkpoint can produce either an R1-style chain-of-thought before the answer or a fast direct response. Qwen3-235B leads several open-weight benchmarks including AIME 2025, LiveCodeBench, ArenaHard and BFCL agent-eval as of release.

Parameters
235B total, 22B active per token
Context
262.144 tokens

Mogelijkheden

  • 235B-parameter MoE with 22B active per token (128 experts, 8 selected)
  • Hybrid thinking mode toggle (CoT or direct answer in one checkpoint)
  • Pretrained on ~36T tokens across 119 languages
  • 256K context window with YaRN scaling
  • Apache 2.0 license, fully commercial
  • Top open-weight scores on AIME 2025, LiveCodeBench, ArenaHard, BFCL
  • Function calling, MCP server support, parallel tool calls
  • Specialised siblings: Qwen3-Coder, Qwen3-Math, Qwen3-VL
  • Compatible with vLLM, SGLang, llama.cpp, Ollama, MLX, HuggingFace
  • Broad multilingual coverage with strong Chinese/English/Japanese/Korean performance
  • Best for: open-weight reasoning, agentic workloads, multilingual chat, on-prem enterprise.

Training & licentie

Pretrained on approximately 36 trillion tokens covering 119 languages with strong Chinese and English emphasis, code repositories and scientific content. Post-training uses a four-stage pipeline: long-CoT cold start, reasoning RL, mixed SFT integrating thinking/non-thinking modes, and a final general-purpose RL stage.

Licentie: Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).

Veiligheidstests: Standard SFT+DPO safety alignment plus general RL safety stage. Filters Chinese politically sensitive topics. Limited third-party safety evaluations published.

Bekende beperkingen

  • Filters Chinese political topics
  • Large memory footprint requires multi-GPU inference for FP16
  • Vision requires separate Qwen3-VL checkpoint
  • Knowledge cutoff approximately early 2025
  • Long context >128K degrades on some recall tasks
04

Prijzen

Momenteel niet beschikbaar: dit model is gedeactiveerd. Er is momenteel geen prijs voor dit model, dus het kan niet worden uitgevoerd.

05

API

Roep Qwen 3 235B Instruct aan met je Railwail API-sleutel. Gebruik deze model-ID in het verzoek:

Momenteel niet beschikbaar

Het model heeft geen geverifieerde prijs of is gedeactiveerd; API-aanroepen worden geweigerd.

06

Specificaties

Model-ID
qwen-3-235b
Ontwikkelaar
Alibaba / Qwen
Categorie
Tekst & chat
Invoer
Tekst
Uitvoer
Tekst
Contextvenster
131.072 tokens
Max. uitvoer
16.384 tokens
Modelgrootte
235B total, 22B active per token
Licentie
Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).
Catalogusitem bijgewerkt
25 juni 2026

Invoerparameters

Invoeren en instellingen uit het invoerschema van het model. Het voorbeeld in de API-sectie toont welke daarvan de API accepteert.

  • promptVerplicht

    User message

    Type: Tekst
    Standaard: –
    Toegestane waarden: tot 16.000 tekens
  • top_p
    Type: Getal
    Standaard: 1
    Toegestane waarden: 0 tot 1
  • stream
    Type: Ja/Nee
    Standaard: false
    Toegestane waarden: –
  • max_tokens
    Type: Geheel getal
    Standaard: 2048
    Toegestane waarden: 1 tot 16.384
  • temperature
    Type: Getal
    Standaard: 0.7
    Toegestane waarden: 0 tot 2
  • system_prompt

    Optional system instruction

    Type: Tekst
    Standaard: –
    Toegestane waarden: tot 8.000 tekens

Tags

  • qwen
  • alibaba
  • moe
  • open-weights
  • flagship
07

Gebruiksscenario's

Waarvoor het wordt gebruikt

  • Open-weight reasoning workloads
  • Agentic tool-using applications
  • Multilingual chat across 100+ languages
  • On-prem enterprise deployments
  • Fine-tuning base for vertical models
  • Apache-2.0-required commercial products
08

Veelgestelde vragen

Wat is Qwen 3 235B Instruct?

Qwen 3 235B Instruct is een model van Alibaba / Qwen in de categorie Tekst & chat. Het staat in de Railwail-catalogus, maar kan momenteel niet worden uitgevoerd.

Hoeveel kost Qwen 3 235B Instruct op Railwail?

Qwen 3 235B Instruct kan momenteel niet op Railwail worden uitgevoerd, dus er is geen huidige prijs. Beschikbare alternatieven met prijzen staan verderop op deze pagina.

Wat is het contextvenster van Qwen 3 235B Instruct?

Het contextvenster van Qwen 3 235B Instruct bevat 131.072 tokens. Een antwoord kan tot 16.384 tokens lang zijn.

Hoe snel is Qwen 3 235B Instruct?

Er zijn nog niet genoeg gemeten runs van Qwen 3 235B Instruct op Railwail om een uitvoeringstijd op te geven. Dit hangt af van de invoer, de instellingen en de belasting bij de provider.

Is Qwen 3 235B Instruct beter dan Claude Fable 5.1?

Dat hangt van de taak af. Qwen 3 235B Instruct (Alibaba / Qwen) en Claude Fable 5.1 (Anthropic) zijn beide modellen in de categorie Tekst & chat. De vergelijkingspagina toont hun prijzen en specificaties naast elkaar.

Qwen 3 235B Instruct en Claude Fable 5.1 vergelijken

Kan ik Qwen 3 235B Instruct nu gebruiken?

Momenteel niet beschikbaar: dit model is gedeactiveerd. De pagina blijft online; beschikbare alternatieven uit dezelfde categorie staan verderop.

Alle modellen via één API

Één API-sleutel voor elk model op Railwail. Gebruik wordt afgerekend via vooraf gekochte credits, 1 credit = US$ 0,01.