Qwen 3 235B Instruct

Text & ChatNicht verfügbar
von Alibaba / QwenModell-ID: qwen-3-235b

Alibaba's Qwen 3 flagship MoE: 235B total / 22B active. Strong reasoning and tool use, open-weights.

Status
Nicht verfügbar
Kontext
131.072 Token
Max. Ausgabe
16.384 Token
Eingabe → Ausgabe
Text → Text
Entwickler
Alibaba / Qwen
Aktualisiert
25. Juni 2026

Qwen 3 235B Instruct ist derzeit nicht verfügbar

Derzeit nicht verfügbar: Dieses Modell ist deaktiviert.

Die Angaben auf dieser Seite kannst du weiter nachlesen. Mit einer der verfügbaren Alternativen unten kannst du sofort ein vergleichbares Modell nutzen.

Zu den Alternativen
01

Vergleichbare Modelle

Alle dieser Kategorie
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    12,00 $/1 Mio. In

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    6,00 $/1 Mio. In

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    4,80 $/1 Mio. In

02

Playground

Qwen 3 235B Instruct ausprobieren

Chat

Derzeit nicht verfügbar

Derzeit nicht verfügbar: Dieses Modell ist deaktiviert.

Der Playground ist deaktiviert. Vergleichbare Modelle findest du in derselben Kategorie: Alternativen ansehen

Qwen 3 235B Instruct ausprobieren

Schick eine Nachricht. Die Antwort kommt vollständig, sobald das Modell fertig ist (ohne Streaming).

System-Prompt
Max. Antwortlänge (Token)

Dieser Lauf

Kein Preis – derzeit nicht verfügbar.

Neu hier?

5 Gratis-Credits (0,05 $) bei Anmeldung mit Google

Nutzbar 24 Stunden nach der Anmeldung, bis zu 5 Läufe pro Tag und höchstens 2 Credits je Lauf. Andere Anmeldearten starten ohne Guthaben.

03

Über Qwen 3 235B Instruct

Kurz gesagtStand: 25. Juni 2026

Qwen 3 235B Instruct ist ein Modell von Alibaba / Qwen aus der Kategorie Text & Chat. Über Railwail ist Qwen 3 235B Instruct derzeit nicht verfügbar. Das Kontextfenster umfasst 131.072 Token, eine Antwort bis zu 16.384 Token.

Hintergrund

Über Alibaba Cloud (Qwen team)

Gegründet 2009 · Hangzhou, China

The Qwen team inside Alibaba Cloud has shipped one of the most prolific open-weight model lines in industry, starting with Qwen-7B (Aug 2023, first Chinese open-weight foundation model from a hyperscaler), through Qwen 1.5 (Feb 2024), Qwen2 (Jun 2024), Qwen2.5 (Sep 2024, sizes 0.5B-72B plus Coder/Math/VL), Qwen2.5-Max closed-weight flagship (Jan 2025) and the Qwen3 family in April-May 2025. Qwen3 introduced a uniform 'hybrid thinking' approach across the family, letting developers toggle between fast direct answers and long chain-of-thought reasoning in the same checkpoint. The 235B-A22B MoE flagship anchors the family alongside dense variants from 0.6B to 32B. The team is led by Junyang Lin and has published over a dozen technical reports. Models ship under the Apache 2.0 license starting with Qwen3, a substantial liberalisation over the earlier Tongyi Qianwen LICENSE. Alibaba Cloud, founded in 2009, is the largest cloud provider in China and hosts the Qwen models on its Model Studio service while also distributing weights freely on HuggingFace, ModelScope and GitHub.

Alibaba Cloud (Qwen team) besuchen

Architektur

Sparse Mixture-of-Experts Transformer (hybrid thinking)

Qwen3-235B-A22B is the flagship Mixture-of-Experts model of the Qwen3 family, released by Alibaba's Qwen team on 29 April 2025 with weights under Apache 2.0. The architecture is a Sparse MoE Transformer with 235 billion total parameters and 22 billion active per token (128 experts, 8 selected per token), 94 layers, and Grouped Query Attention (64 query heads / 4 KV heads). It supports a 128K native context window extended to 256K via YaRN scaling. The model was pretrained on approximately 36 trillion tokens spanning 119 languages with strong Chinese, English and code coverage, more than doubling the 18T-token corpus used for Qwen2.5. Pretraining was performed in three stages with progressively longer context lengths and improved data filtering. Post-training applied a four-stage pipeline: long-CoT cold start, RL on reasoning tasks, integration of thinking and non-thinking modes via mixed SFT, and a final general-purpose RL stage. The resulting model exposes hybrid thinking, toggled via the chat template, where the same checkpoint can produce either an R1-style chain-of-thought before the answer or a fast direct response. Qwen3-235B leads several open-weight benchmarks including AIME 2025, LiveCodeBench, ArenaHard and BFCL agent-eval as of release.

Parameter
235B total, 22B active per token
Kontext
262.144 Token

Fähigkeiten

  • 235B-parameter MoE with 22B active per token (128 experts, 8 selected)
  • Hybrid thinking mode toggle (CoT or direct answer in one checkpoint)
  • Pretrained on ~36T tokens across 119 languages
  • 256K context window with YaRN scaling
  • Apache 2.0 license, fully commercial
  • Top open-weight scores on AIME 2025, LiveCodeBench, ArenaHard, BFCL
  • Function calling, MCP server support, parallel tool calls
  • Specialised siblings: Qwen3-Coder, Qwen3-Math, Qwen3-VL
  • Compatible with vLLM, SGLang, llama.cpp, Ollama, MLX, HuggingFace
  • Broad multilingual coverage with strong Chinese/English/Japanese/Korean performance
  • Best for: open-weight reasoning, agentic workloads, multilingual chat, on-prem enterprise.

Training & Lizenz

Pretrained on approximately 36 trillion tokens covering 119 languages with strong Chinese and English emphasis, code repositories and scientific content. Post-training uses a four-stage pipeline: long-CoT cold start, reasoning RL, mixed SFT integrating thinking/non-thinking modes, and a final general-purpose RL stage.

Lizenz: Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).

Sicherheitstests: Standard SFT+DPO safety alignment plus general RL safety stage. Filters Chinese politically sensitive topics. Limited third-party safety evaluations published.

Bekannte Grenzen

  • Filters Chinese political topics
  • Large memory footprint requires multi-GPU inference for FP16
  • Vision requires separate Qwen3-VL checkpoint
  • Knowledge cutoff approximately early 2025
  • Long context >128K degrades on some recall tasks
04

Preise

Derzeit nicht verfügbar: Dieses Modell ist deaktiviert. Für dieses Modell gibt es derzeit keinen Preis, deshalb lässt es sich nicht ausführen.

05

API

Rufe Qwen 3 235B Instruct mit deinem Railwail-API-Schlüssel auf. Diese Modell-ID gehört in die Anfrage:

Derzeit nicht verfügbar

Das Modell hat keinen geprüften Preis oder ist deaktiviert; API-Aufrufe werden abgelehnt.

06

Spezifikationen

Modell-ID
qwen-3-235b
Entwickler
Alibaba / Qwen
Kategorie
Text & Chat
Eingabe
Text
Ausgabe
Text
Kontextfenster
131.072 Token
Max. Ausgabe
16.384 Token
Modellgröße
235B total, 22B active per token
Lizenz
Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).
Katalogeintrag aktualisiert
25. Juni 2026

Eingabeparameter

Eingaben und Einstellungen laut Eingabeschema des Modells. Welche davon die API annimmt, zeigt das Beispiel im Abschnitt API.

  • promptPflicht

    User message

    Typ: Text
    Standard:
    Erlaubte Werte: bis 16.000 Zeichen
  • top_p
    Typ: Zahl
    Standard: 1
    Erlaubte Werte: 0 bis 1
  • stream
    Typ: Ja/Nein
    Standard: false
    Erlaubte Werte:
  • max_tokens
    Typ: Ganzzahl
    Standard: 2048
    Erlaubte Werte: 1 bis 16.384
  • temperature
    Typ: Zahl
    Standard: 0.7
    Erlaubte Werte: 0 bis 2
  • system_prompt

    Optional system instruction

    Typ: Text
    Standard:
    Erlaubte Werte: bis 8.000 Zeichen

Schlagwörter

  • qwen
  • alibaba
  • moe
  • open-weights
  • flagship
07

Einsatzgebiete

Wofür es genutzt wird

  • Open-weight reasoning workloads
  • Agentic tool-using applications
  • Multilingual chat across 100+ languages
  • On-prem enterprise deployments
  • Fine-tuning base for vertical models
  • Apache-2.0-required commercial products
08

Häufige Fragen

Was ist Qwen 3 235B Instruct?

Qwen 3 235B Instruct ist ein Modell von Alibaba / Qwen aus der Kategorie Text & Chat. Es steht im Railwail-Katalog, lässt sich derzeit aber nicht ausführen.

Was kostet Qwen 3 235B Instruct bei Railwail?

Qwen 3 235B Instruct lässt sich über Railwail derzeit nicht ausführen, deshalb gibt es keinen aktuellen Preis. Verfügbare Alternativen mit Preisen stehen weiter unten auf dieser Seite.

Wie groß ist das Kontextfenster von Qwen 3 235B Instruct?

Das Kontextfenster von Qwen 3 235B Instruct umfasst 131.072 Token. Eine Antwort kann bis zu 16.384 Token lang sein.

Wie schnell ist Qwen 3 235B Instruct?

Für Qwen 3 235B Instruct gibt es bei Railwail noch zu wenige gemessene Läufe, um eine Laufzeit anzugeben. Sie hängt von der Eingabe, den Einstellungen und der Auslastung beim Anbieter ab.

Ist Qwen 3 235B Instruct besser als Claude Fable 5.1?

Das hängt von der Aufgabe ab. Qwen 3 235B Instruct (Alibaba / Qwen) und Claude Fable 5.1 (Anthropic) sind beide Modelle aus der Kategorie Text & Chat. Die Vergleichsseite zeigt Preise und Spezifikationen nebeneinander.

Qwen 3 235B Instruct und Claude Fable 5.1 vergleichen

Kann ich Qwen 3 235B Instruct gerade nutzen?

Derzeit nicht verfügbar: Dieses Modell ist deaktiviert. Die Seite bleibt online; verfügbare Alternativen aus derselben Kategorie stehen weiter unten.

Alle Modelle über eine API

Ein API-Schlüssel für alle Modelle auf Railwail. Abgerechnet wird über vorab gekaufte Credits, 1 Credit = 0,01 $.