Qwen 3 235B Instruct

Tekst i chatNiedostępne
od Alibaba / QwenID modelu: qwen-3-235b

Alibaba's Qwen 3 flagship MoE: 235B total / 22B active. Strong reasoning and tool use, open-weights.

Status
Niedostępne
Kontekst
131 072 tokenów
Maks. wyjście
16 384 tokenów
Wejście → Wyjście
Tekst → Tekst
Deweloper
Alibaba / Qwen
Zaktualizowano
25 czerwca 2026

Qwen 3 235B Instruct jest obecnie niedostępny

Obecnie niedostępne: ten model został dezaktywowany.

Możesz nadal przeczytać szczegóły na tej stronie. Wybierz jedną z dostępnych alternatyw poniżej, aby od razu uruchomić porównywalny model.

Przejdź do alternatyw
01

Porównywalne modele

Wszystkie w tej kategorii
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    12,00 USD/1M wej.

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    6,00 USD/1M wej.

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    4,80 USD/1M wej.

02

Playground

Spróbuj Qwen 3 235B Instruct

Chat

Niedostępny

Obecnie niedostępne: ten model został dezaktywowany.

Plac zabaw jest wyłączony. Porównywalne modele znajdziesz w tej samej kategorii: Przeglądaj alternatywy

Spróbuj Qwen 3 235B Instruct

Wyślij wiadomość. Odpowiedź pojawi się w całości, gdy model będzie gotowy (bez streamingu).

Prompt systemowy
Maks. długość odpowiedzi (tokeny)

To uruchomienie

Brak ceny – obecnie niedostępne.

Nowy tutaj?

5 darmowych kredytów (0,05 USD) po zarejestrowaniu się przez Google

Dostępne 24 godzin po rejestracji, do 5 uruchomień dziennie i maksymalnie 2 kredytów na uruchomienie. Inne metody logowania uruchamiają się bez kredytów.

03

O Qwen 3 235B Instruct

Krótko mówiącStan na 25 czerwca 2026

Qwen 3 235B Instruct to model opracowany przez Alibaba / Qwen w kategorii Tekst i chat. Qwen 3 235B Instruct nie jest obecnie dostępny w serwisie Railwail. Okno kontekstu zawiera 131 072 tokenów, a jedna odpowiedź może mieć do 16 384 tokenów.

Tło

O Alibaba Cloud (Qwen team)

Założona 2009 · Hangzhou, China

The Qwen team inside Alibaba Cloud has shipped one of the most prolific open-weight model lines in industry, starting with Qwen-7B (Aug 2023, first Chinese open-weight foundation model from a hyperscaler), through Qwen 1.5 (Feb 2024), Qwen2 (Jun 2024), Qwen2.5 (Sep 2024, sizes 0.5B-72B plus Coder/Math/VL), Qwen2.5-Max closed-weight flagship (Jan 2025) and the Qwen3 family in April-May 2025. Qwen3 introduced a uniform 'hybrid thinking' approach across the family, letting developers toggle between fast direct answers and long chain-of-thought reasoning in the same checkpoint. The 235B-A22B MoE flagship anchors the family alongside dense variants from 0.6B to 32B. The team is led by Junyang Lin and has published over a dozen technical reports. Models ship under the Apache 2.0 license starting with Qwen3, a substantial liberalisation over the earlier Tongyi Qianwen LICENSE. Alibaba Cloud, founded in 2009, is the largest cloud provider in China and hosts the Qwen models on its Model Studio service while also distributing weights freely on HuggingFace, ModelScope and GitHub.

Odwiedź Alibaba Cloud (Qwen team)

Architektura

Sparse Mixture-of-Experts Transformer (hybrid thinking)

Qwen3-235B-A22B is the flagship Mixture-of-Experts model of the Qwen3 family, released by Alibaba's Qwen team on 29 April 2025 with weights under Apache 2.0. The architecture is a Sparse MoE Transformer with 235 billion total parameters and 22 billion active per token (128 experts, 8 selected per token), 94 layers, and Grouped Query Attention (64 query heads / 4 KV heads). It supports a 128K native context window extended to 256K via YaRN scaling. The model was pretrained on approximately 36 trillion tokens spanning 119 languages with strong Chinese, English and code coverage, more than doubling the 18T-token corpus used for Qwen2.5. Pretraining was performed in three stages with progressively longer context lengths and improved data filtering. Post-training applied a four-stage pipeline: long-CoT cold start, RL on reasoning tasks, integration of thinking and non-thinking modes via mixed SFT, and a final general-purpose RL stage. The resulting model exposes hybrid thinking, toggled via the chat template, where the same checkpoint can produce either an R1-style chain-of-thought before the answer or a fast direct response. Qwen3-235B leads several open-weight benchmarks including AIME 2025, LiveCodeBench, ArenaHard and BFCL agent-eval as of release.

Parametry
235B total, 22B active per token
Kontekst
262 144 tokenów

Możliwości

  • 235B-parameter MoE with 22B active per token (128 experts, 8 selected)
  • Hybrid thinking mode toggle (CoT or direct answer in one checkpoint)
  • Pretrained on ~36T tokens across 119 languages
  • 256K context window with YaRN scaling
  • Apache 2.0 license, fully commercial
  • Top open-weight scores on AIME 2025, LiveCodeBench, ArenaHard, BFCL
  • Function calling, MCP server support, parallel tool calls
  • Specialised siblings: Qwen3-Coder, Qwen3-Math, Qwen3-VL
  • Compatible with vLLM, SGLang, llama.cpp, Ollama, MLX, HuggingFace
  • Broad multilingual coverage with strong Chinese/English/Japanese/Korean performance
  • Best for: open-weight reasoning, agentic workloads, multilingual chat, on-prem enterprise.

Trening i licencja

Pretrained on approximately 36 trillion tokens covering 119 languages with strong Chinese and English emphasis, code repositories and scientific content. Post-training uses a four-stage pipeline: long-CoT cold start, reasoning RL, mixed SFT integrating thinking/non-thinking modes, and a final general-purpose RL stage.

Licencja: Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).

Testy bezpieczeństwa: Standard SFT+DPO safety alignment plus general RL safety stage. Filters Chinese politically sensitive topics. Limited third-party safety evaluations published.

Znane ograniczenia

  • Filters Chinese political topics
  • Large memory footprint requires multi-GPU inference for FP16
  • Vision requires separate Qwen3-VL checkpoint
  • Knowledge cutoff approximately early 2025
  • Long context >128K degrades on some recall tasks
04

Ceny

Obecnie niedostępne: ten model został dezaktywowany. Dla tego modelu nie ma ceny w tej chwili, dlatego nie można go uruchomić.

05

API

Wywołaj Qwen 3 235B Instruct za pomocą klucza API Railwail. Użyj tego ID modelu w żądaniu:

Obecnie niedostępne

Model nie ma zweryfikowanej ceny lub jest wyłączony; wywołania API są odrzucane.

06

Specyfikacje

ID modelu
qwen-3-235b
Deweloper
Alibaba / Qwen
Kategoria
Tekst i chat
Wejście
Tekst
Wyjście
Tekst
Okno kontekstu
131 072 tokenów
Maks. wyjście
16 384 tokenów
Rozmiar modelu
235B total, 22B active per token
Licencja
Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).
Wpis w katalogu zaktualizowany
25 czerwca 2026

Parametry wejściowe

Dane wejściowe i ustawienia ze schematu wejściowego modelu. Przykład w sekcji API pokazuje, które z nich API akceptuje.

  • promptwymagane

    User message

    Typ: Tekst
    Domyślnie:
    Dozwolone wartości: do 16 000 znaków
  • top_p
    Typ: Liczba
    Domyślnie: 1
    Dozwolone wartości: 0 do 1
  • stream
    Typ: Tak/Nie
    Domyślnie: false
    Dozwolone wartości:
  • max_tokens
    Typ: Liczba całkowita
    Domyślnie: 2048
    Dozwolone wartości: 1 do 16 384
  • temperature
    Typ: Liczba
    Domyślnie: 0.7
    Dozwolone wartości: 0 do 2
  • system_prompt

    Optional system instruction

    Typ: Tekst
    Domyślnie:
    Dozwolone wartości: do 8000 znaków

Tagi

  • qwen
  • alibaba
  • moe
  • open-weights
  • flagship
07

Przypadki użycia

Do czego się go używa

  • Open-weight reasoning workloads
  • Agentic tool-using applications
  • Multilingual chat across 100+ languages
  • On-prem enterprise deployments
  • Fine-tuning base for vertical models
  • Apache-2.0-required commercial products
08

Często zadawane pytania

Co to jest Qwen 3 235B Instruct?

Qwen 3 235B Instruct to model opracowany przez Alibaba / Qwen w kategorii Tekst i chat. Jest wymieniony w katalogu Railwail, ale nie może być uruchomiony w tej chwili.

Ile kosztuje Qwen 3 235B Instruct w serwisie Railwail?

Qwen 3 235B Instruct nie może być uruchomiony w serwisie Railwail w tej chwili, dlatego nie ma aktualnej ceny. Dostępne alternatywy z cenami są wymienione poniżej na tej stronie.

Jakie jest okno kontekstu Qwen 3 235B Instruct?

Okno kontekstu Qwen 3 235B Instruct zawiera 131 072 tokenów. Jedna odpowiedź może mieć do 16 384 tokenów.

Jak szybki jest Qwen 3 235B Instruct?

Dla Qwen 3 235B Instruct jest jeszcze zbyt mało zmierzonych przebiegów w serwisie Railwail, aby podać czas przebiegu. Zależy to od wejścia, ustawień i obciążenia u dostawcy.

Czy Qwen 3 235B Instruct jest lepszy niż Claude Fable 5.1?

To zależy od zadania. Qwen 3 235B Instruct (Alibaba / Qwen) i Claude Fable 5.1 (Anthropic) to oba modele z kategorii Tekst i chat. Strona porównania pokazuje ich ceny i specyfikacje obok siebie.

Porównaj Qwen 3 235B Instruct i Claude Fable 5.1

Czy mogę używać Qwen 3 235B Instruct teraz?

Obecnie niedostępne: ten model został dezaktywowany. Strona pozostaje online; dostępne alternatywy z tej samej kategorii są wymienione poniżej.

Wszystkie modele przez jedno API

Jeden klucz API dla każdego modelu na Railwail. Opłaty pobierane są z przedpłaconych kredytów, 1 kredyt = 0,01 USD.