Qwen 3 235B Instruct

Текст и чатНедоступно
от Alibaba / QwenID модели: qwen-3-235b

Alibaba's Qwen 3 flagship MoE: 235B total / 22B active. Strong reasoning and tool use, open-weights.

Статус
Недоступно
Контекст
131 072 токенов
Макс. выход
16 384 токенов
Вход → выход
Текст → Текст
Разработчик
Alibaba / Qwen
Обновлено
25 июня 2026 г.

Qwen 3 235B Instruct сейчас недоступна

Недоступно: эта модель деактивирована.

Вы можете прочитать подробности на этой странице. Выберите одну из доступных альтернатив ниже, чтобы сразу запустить сравнимую модель.

К альтернативам
01

Сравнимые модели

Все в этой категории
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    12,00 $/1M вход

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    6,00 $/1M вход

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    4,80 $/1M вход

02

Playground

Попробовать Qwen 3 235B Instruct

Чат

Сейчас недоступна

Недоступно: эта модель деактивирована.

Playground отключен. Сравнимые модели найдёте в той же категории: Посмотреть альтернативы

Попробовать Qwen 3 235B Instruct

Отправьте сообщение. Ответ придёт полностью, когда модель закончит работу (без потоковой передачи).

Системный промпт
Макс. длина ответа (токены)

Этот запуск

Нет цены – сейчас недоступно.

Впервые здесь?

5 бесплатных кредитов (0,05 $) при регистрации через Google

Доступно 24 часов после регистрации, до 5 запусков в день и максимум 2 кредитов за запуск. Другие способы входа начинают без кредитов.

03

О Qwen 3 235B Instruct

КороткоПо состоянию на 25 июня 2026 г.

Qwen 3 235B Instruct — это модель от Alibaba / Qwen в категории Текст и чат. Qwen 3 235B Instruct сейчас недоступна на Railwail. Контекстное окно содержит 131 072 токенов, ответ может быть до 16 384 токенов.

Фон

О Alibaba Cloud (Qwen team)

Основана в 2009 · Hangzhou, China

The Qwen team inside Alibaba Cloud has shipped one of the most prolific open-weight model lines in industry, starting with Qwen-7B (Aug 2023, first Chinese open-weight foundation model from a hyperscaler), through Qwen 1.5 (Feb 2024), Qwen2 (Jun 2024), Qwen2.5 (Sep 2024, sizes 0.5B-72B plus Coder/Math/VL), Qwen2.5-Max closed-weight flagship (Jan 2025) and the Qwen3 family in April-May 2025. Qwen3 introduced a uniform 'hybrid thinking' approach across the family, letting developers toggle between fast direct answers and long chain-of-thought reasoning in the same checkpoint. The 235B-A22B MoE flagship anchors the family alongside dense variants from 0.6B to 32B. The team is led by Junyang Lin and has published over a dozen technical reports. Models ship under the Apache 2.0 license starting with Qwen3, a substantial liberalisation over the earlier Tongyi Qianwen LICENSE. Alibaba Cloud, founded in 2009, is the largest cloud provider in China and hosts the Qwen models on its Model Studio service while also distributing weights freely on HuggingFace, ModelScope and GitHub.

Посетить Alibaba Cloud (Qwen team)

Архитектура

Sparse Mixture-of-Experts Transformer (hybrid thinking)

Qwen3-235B-A22B is the flagship Mixture-of-Experts model of the Qwen3 family, released by Alibaba's Qwen team on 29 April 2025 with weights under Apache 2.0. The architecture is a Sparse MoE Transformer with 235 billion total parameters and 22 billion active per token (128 experts, 8 selected per token), 94 layers, and Grouped Query Attention (64 query heads / 4 KV heads). It supports a 128K native context window extended to 256K via YaRN scaling. The model was pretrained on approximately 36 trillion tokens spanning 119 languages with strong Chinese, English and code coverage, more than doubling the 18T-token corpus used for Qwen2.5. Pretraining was performed in three stages with progressively longer context lengths and improved data filtering. Post-training applied a four-stage pipeline: long-CoT cold start, RL on reasoning tasks, integration of thinking and non-thinking modes via mixed SFT, and a final general-purpose RL stage. The resulting model exposes hybrid thinking, toggled via the chat template, where the same checkpoint can produce either an R1-style chain-of-thought before the answer or a fast direct response. Qwen3-235B leads several open-weight benchmarks including AIME 2025, LiveCodeBench, ArenaHard and BFCL agent-eval as of release.

Параметры
235B total, 22B active per token
Контекст
262 144 токенов

Возможности

  • 235B-parameter MoE with 22B active per token (128 experts, 8 selected)
  • Hybrid thinking mode toggle (CoT or direct answer in one checkpoint)
  • Pretrained on ~36T tokens across 119 languages
  • 256K context window with YaRN scaling
  • Apache 2.0 license, fully commercial
  • Top open-weight scores on AIME 2025, LiveCodeBench, ArenaHard, BFCL
  • Function calling, MCP server support, parallel tool calls
  • Specialised siblings: Qwen3-Coder, Qwen3-Math, Qwen3-VL
  • Compatible with vLLM, SGLang, llama.cpp, Ollama, MLX, HuggingFace
  • Broad multilingual coverage with strong Chinese/English/Japanese/Korean performance
  • Best for: open-weight reasoning, agentic workloads, multilingual chat, on-prem enterprise.

Обучение и лицензия

Pretrained on approximately 36 trillion tokens covering 119 languages with strong Chinese and English emphasis, code repositories and scientific content. Post-training uses a four-stage pipeline: long-CoT cold start, reasoning RL, mixed SFT integrating thinking/non-thinking modes, and a final general-purpose RL stage.

Лицензия: Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).

Тестирование безопасности: Standard SFT+DPO safety alignment plus general RL safety stage. Filters Chinese politically sensitive topics. Limited third-party safety evaluations published.

Известные ограничения

  • Filters Chinese political topics
  • Large memory footprint requires multi-GPU inference for FP16
  • Vision requires separate Qwen3-VL checkpoint
  • Knowledge cutoff approximately early 2025
  • Long context >128K degrades on some recall tasks
04

Цены

Недоступно: эта модель деактивирована. На данный момент для этой модели нет цены, поэтому её нельзя запустить.

05

API

Вызовите Qwen 3 235B Instruct с помощью вашего API-ключа Railwail. Используйте этот ID модели в запросе:

Временно недоступно

Модель не имеет проверенную цену или деактивирована; вызовы API отклоняются.

06

Спецификации

ID модели
qwen-3-235b
Разработчик
Alibaba / Qwen
Категория
Текст и чат
Входные данные
Текст
Выходные данные
Текст
Контекстное окно
131 072 токенов
Макс. выходные данные
16 384 токенов
Размер модели
235B total, 22B active per token
Лицензия
Apache 2.0. Open weights, fully commercial use permitted including for >100M MAU products (a notable liberalisation versus the earlier Tongyi Qianwen LICENSE).
Запись в каталоге обновлена
25 июня 2026 г.

Входные параметры

Входные данные и параметры из схемы входных данных модели. Пример в разделе API показывает, какие из них принимает API.

  • promptобязательно

    User message

    Тип: Текст
    По умолчанию:
    Допустимые значения: до 16 000 символов
  • top_p
    Тип: Число
    По умолчанию: 1
    Допустимые значения: от 0 до 1
  • stream
    Тип: Да/Нет
    По умолчанию: false
    Допустимые значения:
  • max_tokens
    Тип: Целое число
    По умолчанию: 2048
    Допустимые значения: от 1 до 16 384
  • temperature
    Тип: Число
    По умолчанию: 0.7
    Допустимые значения: от 0 до 2
  • system_prompt

    Optional system instruction

    Тип: Текст
    По умолчанию:
    Допустимые значения: до 8 000 символов

Теги

  • qwen
  • alibaba
  • moe
  • open-weights
  • flagship
07

Применение

Для чего это используется

  • Open-weight reasoning workloads
  • Agentic tool-using applications
  • Multilingual chat across 100+ languages
  • On-prem enterprise deployments
  • Fine-tuning base for vertical models
  • Apache-2.0-required commercial products
08

Часто задаваемые вопросы

Что такое Qwen 3 235B Instruct?

Qwen 3 235B Instruct — модель от Alibaba / Qwen в категории Текст и чат. Модель указана в каталоге Railwail, но сейчас её нельзя запустить.

Сколько стоит Qwen 3 235B Instruct на Railwail?

Qwen 3 235B Instruct сейчас нельзя запустить на Railwail, поэтому текущей цены нет. Доступные альтернативы с ценами указаны ниже на этой странице.

Какой размер контекстного окна у Qwen 3 235B Instruct?

Контекстное окно Qwen 3 235B Instruct содержит 131 072 токенов. Ответ может быть до 16 384 токенов.

Насколько быстра Qwen 3 235B Instruct?

На Railwail пока недостаточно измеренных запусков Qwen 3 235B Instruct, чтобы указать время выполнения. Оно зависит от входных данных, параметров и нагрузки на провайдера.

Qwen 3 235B Instruct лучше, чем Claude Fable 5.1?

Это зависит от задачи. Qwen 3 235B Instruct (Alibaba / Qwen) и Claude Fable 5.1 (Anthropic) — обе модели в категории Текст и чат. На странице сравнения показаны их цены и характеристики рядом.

Сравнить Qwen 3 235B Instruct и Claude Fable 5.1

Могу ли я использовать Qwen 3 235B Instruct прямо сейчас?

Недоступно: эта модель деактивирована. Страница остаётся в сети; доступные альтернативы из той же категории указаны ниже.

Все модели через один API

Один API-ключ для всех моделей на Railwail. Использование оплачивается из предоплаченных кредитов, 1 кредит = 0,01 $.