Qwen 3 235B Instruct

Text și chatIndisponibil
de Alibaba / QwenID model: qwen-3-235b

Alibaba's Qwen 3 flagship MoE: 235B total / 22B active. Strong reasoning and tool use, open-weights.

Status
Indisponibil
Context
131.072 tokeni
Ieșire max.
16.384 tokeni
Intrare → ieșire
Text → Text
Dezvoltator
Alibaba / Qwen
Actualizat
25 iunie 2026

Qwen 3 235B Instruct nu este disponibil în acest moment

Indisponibil în prezent: acest model a fost dezactivat.

Poți citi în continuare detaliile pe această pagină. Alege una dintre alternativele disponibile de mai jos pentru a rula imediat un model comparabil.

Mergi la alternative
01
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    12,00 USD/1M in

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    6,00 USD/1M in

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    4,80 USD/1M in

02

Playground

Încearcă Qwen 3 235B Instruct

Chat

Indisponibil în prezent

Indisponibil în prezent: acest model a fost dezactivat.

Playground-ul este dezactivat. Modele comparabile găsești în aceeași categorie: Vezi alternativele

Încearcă Qwen 3 235B Instruct

Trimite un mesaj. Răspunsul sosește complet când modelul termină (fără streaming).

Prompt de sistem
Lungimea max. a răspunsului (tokeni)

Această rulare

Fără preț – momentan indisponibil.

Nou aici?

5 credite gratuite (0,05 USD) când te înregistrezi cu Google

Utilizabil 24 ore după înregistrare, până la 5 rulări pe zi și maximum 2 credite pe rulare. Alte metode de conectare încep fără credite.

03

Despre Qwen 3 235B Instruct

Pe scurtDin 25 iunie 2026

Qwen 3 235B Instruct este un model de Alibaba / Qwen din categoria Text și chat. Qwen 3 235B Instruct nu este disponibil în prezent pe Railwail. Fereastra de context conține 131.072 token-uri, iar un răspuns poate fi lung de până la 16.384 token-uri.

Fundal

Despre Alibaba Cloud (Qwen team)

Fondat 2009 · Hangzhou, China

The Qwen team inside Alibaba Cloud has shipped one of the most prolific open-weight model lines in industry, starting with Qwen-7B (Aug 2023, first Chinese open-weight foundation model from a hyperscaler), through Qwen 1.5 (Feb 2024), Qwen2 (Jun 2024), Qwen2.5 (Sep 2024, sizes 0.5B-72B plus Coder/Math/VL), Qwen2.5-Max closed-weight flagship (Jan 2025) and the Qwen3 family in April-May 2025. Qwen3 introduced a uniform 'hybrid thinking' approach across the family, letting developers toggle between fast direct answers and long chain-of-thought reasoning in the same checkpoint. The 235B-A22B MoE flagship anchors the family alongside dense variants from 0.6B to 32B. The team is led by Junyang Lin and has published over a dozen technical reports. Models ship under the Apache 2.0 license starting with Qwen3, a substantial liberalisation over the earlier Tongyi Qianwen LICENSE. Alibaba Cloud, founded in 2009, is the largest cloud provider in China and hosts the Qwen models on its Model Studio service while also distributing weights freely on HuggingFace, ModelScope and GitHub.

Vizitează Alibaba Cloud (Qwen team)

Arhitectură

Sparse Mixture-of-Experts Transformer (hybrid thinking)

Qwen3-235B-A22B is the flagship Mixture-of-Experts model of the Qwen3 family, released by Alibaba's Qwen team on 29 April 2025 with weights under Apache 2.0. The architecture is a Sparse MoE Transformer with 235 billion total parameters and 22 billion active per token (128 experts, 8 selected per token), 94 layers, and Grouped Query Attention (64 query heads / 4 KV heads). It supports a 128K native context window extended to 256K via YaRN scaling. The model was pretrained on approximately 36 trillion tokens spanning 119 languages with strong Chinese, English and code coverage, more than doubling the 18T-token corpus used for Qwen2.5. Pretraining was performed in three stages with progressively longer context lengths and improved data filtering. Post-training applied a four-stage pipeline: long-CoT cold start, RL on reasoning tasks, integration of thinking and non-thinking modes via mixed SFT, and a final general-purpose RL stage. The resulting model exposes hybrid thinking, toggled via the chat template, where the same checkpoint can produce either an R1-style chain-of-thought before the answer or a fast direct response. Qwen3-235B leads several open-weight benchmarks including AIME 2025, LiveCodeBench, ArenaHard and BFCL agent-eval as of release.

Parametri
235B total, 22B active per token
Context
262.144 tokeni

Capabilități

  • 235B-parameter MoE with 22B active per token (128 experts, 8 selected)
  • Hybrid thinking mode toggle (CoT or direct answer in one checkpoint)
  • Pretrained on ~36T tokens across 119 languages
  • 256K context window with YaRN scaling
  • Apache 2.0 license, fully commercial
  • Top open-weight scores on AIME 2025, LiveCodeBench, ArenaHard, BFCL
  • Function calling, MCP server support, parallel tool calls
  • Specialised siblings: Qwen3-Coder, Qwen3-Math, Qwen3-VL
  • Compatible with vLLM, SGLang, llama.cpp, Ollama, MLX, HuggingFace
  • Broad multilingual coverage with strong Chinese/English/Japanese/Korean performance
  • Best for: open-weight reasoning, agentic workloads, multilingual chat, on-prem enterprise.

Antrenament & licență

Pretrained on approximately 36 trillion tokens covering 119 languages with strong Chinese and English emphasis, code repositories and scientific content. Post-training uses a four-stage pipeline: long-CoT cold start, reasoning RL, mixed SFT integrating thinking/non-thinking modes, and a final general-purpose RL stage.

Licență: Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).

Teste de siguranță: Standard SFT+DPO safety alignment plus general RL safety stage. Filters Chinese politically sensitive topics. Limited third-party safety evaluations published.

Limitări cunoscute

  • Filters Chinese political topics
  • Large memory footprint requires multi-GPU inference for FP16
  • Vision requires separate Qwen3-VL checkpoint
  • Knowledge cutoff approximately early 2025
  • Long context >128K degrades on some recall tasks
04

Prețuri

Indisponibil în prezent: acest model a fost dezactivat. Nu există preț pentru acest model în acest moment, deci nu poate fi executat.

05

API

Apelează Qwen 3 235B Instruct cu cheia ta API Railwail. Folosește acest ID de model în cerere:

Indisponibil în prezent

Modelul nu are un preț verificat sau este dezactivat; apelurile API sunt refuzate.

06

Specificații

ID model
qwen-3-235b
Dezvoltator
Alibaba / Qwen
Categorie
Text și chat
Intrare
Text
Ieșire
Text
Fereastră de context
131.072 tokeni
Ieșire max.
16.384 tokeni
Dimensiune model
235B total, 22B active per token
Licență
Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).
Intrare catalog actualizată
25 iunie 2026

Parametri de intrare

Intrări și setări din schema de intrare a modelului. Exemplul din secțiunea API arată care dintre ele acceptă API-ul.

  • promptobligatoriu

    User message

    Tip: Text
    Implicit:
    Valori permise: până la 16.000 caractere
  • top_p
    Tip: Număr
    Implicit: 1
    Valori permise: 0 până la 1
  • stream
    Tip: Da/Nu
    Implicit: false
    Valori permise:
  • max_tokens
    Tip: Număr întreg
    Implicit: 2048
    Valori permise: 1 până la 16.384
  • temperature
    Tip: Număr
    Implicit: 0.7
    Valori permise: 0 până la 2
  • system_prompt

    Optional system instruction

    Tip: Text
    Implicit:
    Valori permise: până la 8.000 caractere

Etichete

  • qwen
  • alibaba
  • moe
  • open-weights
  • flagship
07

Cazuri de utilizare

Pentru ce se folosește

  • Open-weight reasoning workloads
  • Agentic tool-using applications
  • Multilingual chat across 100+ languages
  • On-prem enterprise deployments
  • Fine-tuning base for vertical models
  • Apache-2.0-required commercial products
08

Întrebări frecvente

Ce este Qwen 3 235B Instruct?

Qwen 3 235B Instruct este un model de Alibaba / Qwen din categoria Text și chat. Este listat pe Railwail, dar nu poate fi rulat în acest moment.

Cât costă Qwen 3 235B Instruct pe Railwail?

Qwen 3 235B Instruct nu poate fi rulat pe Railwail în acest moment, deci nu există preț curent. Alternativele disponibile cu prețuri sunt listate mai jos pe această pagină.

Care este fereastra de context a Qwen 3 235B Instruct?

Fereastra de context a Qwen 3 235B Instruct conține 131.072 token-uri. Un răspuns poate fi lung de până la 16.384 token-uri.

Cât de rapid este Qwen 3 235B Instruct?

Nu sunt suficiente rulări măsurate ale Qwen 3 235B Instruct pe Railwail încă pentru a indica un timp de rulare. Depinde de intrare, de setări și de sarcina la furnizor.

Este Qwen 3 235B Instruct mai bun decât Claude Fable 5.1?

Depinde de sarcină. Qwen 3 235B Instruct (Alibaba / Qwen) și Claude Fable 5.1 (Anthropic) sunt ambele modele din categoria Text și chat. Pagina de comparație arată prețurile și specificațiile lor una lângă alta.

Compară Qwen 3 235B Instruct și Claude Fable 5.1

Pot folosi Qwen 3 235B Instruct chiar acum?

Indisponibil în prezent: acest model a fost dezactivat. Pagina rămâne online; alternativele disponibile din aceeași categorie sunt listate mai jos.

Toate modelele printr-o singură API

O cheie API pentru fiecare model pe Railwail. Utilizarea se percepe din credite prepay, 1 credit = 0,01 USD.