Qwen 3 235B Instruct

Testo e chatNon disponibile
di Alibaba / QwenID modello: qwen-3-235b

Alibaba's Qwen 3 flagship MoE: 235B total / 22B active. Strong reasoning and tool use, open-weights.

Stato
Non disponibile
Contesto
131.072 token
Max. output
16.384 token
Input → output
Testo → Testo
Sviluppatore
Alibaba / Qwen
Aggiornato
25 giugno 2026

Qwen 3 235B Instruct non è attualmente disponibile

Attualmente non disponibile: questo modello è stato disattivato.

Puoi comunque leggere i dettagli su questa pagina. Scegli una delle alternative disponibili di seguito per eseguire subito un modello comparabile.

Vai alle alternative
01

Modelli comparabili

Tutti in questa categoria
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    12,00 USD/1M in

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    6,00 USD/1M in

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    4,80 USD/1M in

02

Playground

Prova Qwen 3 235B Instruct

Chat

Attualmente non disponibile

Attualmente non disponibile: questo modello è stato disattivato.

Il playground è disabilitato. Trovi modelli comparabili nella stessa categoria: Visualizza alternative

Prova Qwen 3 235B Instruct

Invia un messaggio. La risposta arriva completa quando il modello ha finito (senza streaming).

Prompt di sistema
Lunghezza massima della risposta (token)

Questa esecuzione

Nessun prezzo – attualmente non disponibile.

Nuovo qui?

5 crediti gratuiti (0,05 USD) quando ti iscrivi con Google

Utilizzabile 24 ore dopo l'iscrizione, fino a 5 esecuzioni al giorno e al massimo 2 crediti per esecuzione. Altri metodi di accesso iniziano senza crediti.

03

Informazioni su Qwen 3 235B Instruct

RiassuntoA partire da 25 giugno 2026

Qwen 3 235B Instruct è un modello di Alibaba / Qwen nella categoria Testo e chat. Qwen 3 235B Instruct non è attualmente disponibile su Railwail. La finestra di contesto contiene 131.072 token e una risposta può essere lunga fino a 16.384 token.

Sfondo

Informazioni su Alibaba Cloud (Qwen team)

Fondato 2009 · Hangzhou, China

The Qwen team inside Alibaba Cloud has shipped one of the most prolific open-weight model lines in industry, starting with Qwen-7B (Aug 2023, first Chinese open-weight foundation model from a hyperscaler), through Qwen 1.5 (Feb 2024), Qwen2 (Jun 2024), Qwen2.5 (Sep 2024, sizes 0.5B-72B plus Coder/Math/VL), Qwen2.5-Max closed-weight flagship (Jan 2025) and the Qwen3 family in April-May 2025. Qwen3 introduced a uniform 'hybrid thinking' approach across the family, letting developers toggle between fast direct answers and long chain-of-thought reasoning in the same checkpoint. The 235B-A22B MoE flagship anchors the family alongside dense variants from 0.6B to 32B. The team is led by Junyang Lin and has published over a dozen technical reports. Models ship under the Apache 2.0 license starting with Qwen3, a substantial liberalisation over the earlier Tongyi Qianwen LICENSE. Alibaba Cloud, founded in 2009, is the largest cloud provider in China and hosts the Qwen models on its Model Studio service while also distributing weights freely on HuggingFace, ModelScope and GitHub.

Visita Alibaba Cloud (Qwen team)

Architettura

Sparse Mixture-of-Experts Transformer (hybrid thinking)

Qwen3-235B-A22B is the flagship Mixture-of-Experts model of the Qwen3 family, released by Alibaba's Qwen team on 29 April 2025 with weights under Apache 2.0. The architecture is a Sparse MoE Transformer with 235 billion total parameters and 22 billion active per token (128 experts, 8 selected per token), 94 layers, and Grouped Query Attention (64 query heads / 4 KV heads). It supports a 128K native context window extended to 256K via YaRN scaling. The model was pretrained on approximately 36 trillion tokens spanning 119 languages with strong Chinese, English and code coverage, more than doubling the 18T-token corpus used for Qwen2.5. Pretraining was performed in three stages with progressively longer context lengths and improved data filtering. Post-training applied a four-stage pipeline: long-CoT cold start, RL on reasoning tasks, integration of thinking and non-thinking modes via mixed SFT, and a final general-purpose RL stage. The resulting model exposes hybrid thinking, toggled via the chat template, where the same checkpoint can produce either an R1-style chain-of-thought before the answer or a fast direct response. Qwen3-235B leads several open-weight benchmarks including AIME 2025, LiveCodeBench, ArenaHard and BFCL agent-eval as of release.

Parametri
235B total, 22B active per token
Contesto
262.144 token

Capacità

  • 235B-parameter MoE with 22B active per token (128 experts, 8 selected)
  • Hybrid thinking mode toggle (CoT or direct answer in one checkpoint)
  • Pretrained on ~36T tokens across 119 languages
  • 256K context window with YaRN scaling
  • Apache 2.0 license, fully commercial
  • Top open-weight scores on AIME 2025, LiveCodeBench, ArenaHard, BFCL
  • Function calling, MCP server support, parallel tool calls
  • Specialised siblings: Qwen3-Coder, Qwen3-Math, Qwen3-VL
  • Compatible with vLLM, SGLang, llama.cpp, Ollama, MLX, HuggingFace
  • Broad multilingual coverage with strong Chinese/English/Japanese/Korean performance
  • Best for: open-weight reasoning, agentic workloads, multilingual chat, on-prem enterprise.

Addestramento e licenza

Pretrained on approximately 36 trillion tokens covering 119 languages with strong Chinese and English emphasis, code repositories and scientific content. Post-training uses a four-stage pipeline: long-CoT cold start, reasoning RL, mixed SFT integrating thinking/non-thinking modes, and a final general-purpose RL stage.

Licenza: Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).

Test di sicurezza: Standard SFT+DPO safety alignment plus general RL safety stage. Filters Chinese politically sensitive topics. Limited third-party safety evaluations published.

Limitazioni note

  • Filters Chinese political topics
  • Large memory footprint requires multi-GPU inference for FP16
  • Vision requires separate Qwen3-VL checkpoint
  • Knowledge cutoff approximately early 2025
  • Long context >128K degrades on some recall tasks
04

Prezzi

Attualmente non disponibile: questo modello è stato disattivato. Al momento non c'è un prezzo per questo modello, quindi non può essere eseguito.

05

API

Chiama Qwen 3 235B Instruct con la tua chiave API Railwail. Usa questo ID modello nella richiesta:

Attualmente non disponibile

Il modello non ha un prezzo verificato o è disattivato; le chiamate API vengono rifiutate.

06

Specifiche

ID modello
qwen-3-235b
Sviluppatore
Alibaba / Qwen
Categoria
Testo e chat
Input
Testo
Output
Testo
Finestra di contesto
131.072 token
Output massimo
16.384 token
Dimensione del modello
235B total, 22B active per token
Licenza
Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).
Voce di catalogo aggiornata
25 giugno 2026

Parametri di input

Input e impostazioni dallo schema di input del modello. L'esempio nella sezione API mostra quali di essi l'API accetta.

  • promptObbligatorio

    User message

    Tipo: Testo
    Predefinito: –
    Valori consentiti: fino a 16.000 caratteri
  • top_p
    Tipo: Numero
    Predefinito: 1
    Valori consentiti: 0 a 1
  • stream
    Tipo: Sì/No
    Predefinito: false
    Valori consentiti: –
  • max_tokens
    Tipo: Numero intero
    Predefinito: 2048
    Valori consentiti: 1 a 16.384
  • temperature
    Tipo: Numero
    Predefinito: 0.7
    Valori consentiti: 0 a 2
  • system_prompt

    Optional system instruction

    Tipo: Testo
    Predefinito: –
    Valori consentiti: fino a 8000 caratteri

Etichette

  • qwen
  • alibaba
  • moe
  • open-weights
  • flagship
07

Casi d'uso

A cosa serve

  • Open-weight reasoning workloads
  • Agentic tool-using applications
  • Multilingual chat across 100+ languages
  • On-prem enterprise deployments
  • Fine-tuning base for vertical models
  • Apache-2.0-required commercial products
08

Domande frequenti

Cos'è Qwen 3 235B Instruct?

Qwen 3 235B Instruct è un modello di Alibaba / Qwen nella categoria Testo e chat. È elencato su Railwail ma non può essere eseguito al momento.

Quanto costa Qwen 3 235B Instruct su Railwail?

Qwen 3 235B Instruct non può essere eseguito su Railwail al momento, quindi non c'è un prezzo attuale. Le alternative disponibili con i prezzi sono elencate più in basso in questa pagina.

Qual è la finestra di contesto di Qwen 3 235B Instruct?

La finestra di contesto di Qwen 3 235B Instruct contiene 131.072 token. Una risposta può essere lunga fino a 16.384 token.

Quanto è veloce Qwen 3 235B Instruct?

Non ci sono ancora abbastanza esecuzioni misurate di Qwen 3 235B Instruct su Railwail per indicare un tempo di esecuzione. Dipende dall'input, dalle impostazioni e dal carico presso il provider.

Qwen 3 235B Instruct è migliore di Claude Fable 5.1?

Dipende dall'attività. Qwen 3 235B Instruct (Alibaba / Qwen) e Claude Fable 5.1 (Anthropic) sono entrambi modelli nella categoria Testo e chat. La pagina di confronto mostra i loro prezzi e le specifiche affiancati.

Confronta Qwen 3 235B Instruct e Claude Fable 5.1

Posso usare Qwen 3 235B Instruct adesso?

Attualmente non disponibile: questo modello è stato disattivato. La pagina rimane online; le alternative disponibili della stessa categoria sono elencate più in basso.

Tutti i modelli tramite un'API

Una chiave API per tutti i modelli su Railwail. L'utilizzo viene addebitato da crediti prepagati, 1 credito = 0,01 USD.