Qwen 2.5-Max

Text & chatUnavailable
by OtherModel ID: qwen-2-5-max

Alibaba's flagship pretrained MoE model. Top-tier reasoning and code performance via DashScope API.

Status
Unavailable
Context
32,768 tokens
Max. output
8,192 tokens
Input โ†’ output
Text โ†’ Text
Developer
Other
Updated
September 23, 2026

Qwen 2.5-Max is currently unavailable

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

Go to alternatives
01

Comparable models

All in this category
02

Playground

Try Qwen 2.5-Max

Input & output

Currently unavailable

Currently unavailable.

The playground is disabled. You can find comparable models in the same category: Browse alternatives

Try Qwen 2.5-Max
Output
The answer appears here.

This run

No price โ€“ currently unavailable.

New here?

10 free credits ($0.10) when you sign up with Google

Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits.

03

About Qwen 2.5-Max

TL;DRAs of September 23, 2026

Qwen 2.5-Max is a model by Other in the Text & chat category. Qwen 2.5-Max is currently not available on Railwail. The context window holds 32,768 tokens, and one response can be up to 8,192 tokens long.

Background

About Alibaba Cloud (Qwen team)

Founded 2009 ยท Hangzhou, China

The Qwen team is Alibaba Cloud's large-language-model research unit, building on Alibaba's Damo Academy and Tongyi Lab research dating back to the late 2010s. Alibaba Cloud itself was founded in 2009 and is the largest public cloud provider in China. The Qwen series first appeared in 2023 with the open-weight Qwen-7B and Qwen-72B, followed by Qwen 1.5 (Feb 2024), Qwen2 (Jun 2024), Qwen2.5 (Sep 2024) with sizes from 0.5B to 72B plus Coder, Math and VL specialised siblings, Qwen2.5-Max as the closed-weight flagship (Jan 2025) and the Qwen3 family in 2025 introducing built-in thinking mode. The team led by Junyang Lin has published over a dozen technical reports and powers the Tongyi Qianwen consumer chat product across Alibaba properties. Qwen open weights have become the largest derivative-model family on HuggingFace by download volume, with thousands of community fine-tunes. Qwen2.5-Max is positioned as the closed-source flagship that benchmarks against GPT-4o, Claude 3.5 Sonnet and DeepSeek V3, exclusively available via the Alibaba Cloud API.

Visit Alibaba Cloud (Qwen team)

Architecture

Sparse Mixture-of-Experts Transformer (closed-weight flagship)

Qwen2.5-Max is the closed-weight flagship of the Qwen2.5 family, announced by Alibaba Cloud on 29 January 2025 - one week after DeepSeek V3. It is a Sparse Mixture-of-Experts Transformer; Alibaba has not publicly disclosed total or active parameter counts but confirms MoE architecture. The model was pretrained on more than 20 trillion tokens, surpassing Qwen2.5-72B's 18T-token corpus, with continued heavy emphasis on Chinese, English and code. Post-training combines supervised fine-tuning on Alibaba's curated multi-domain instruction set with Reinforcement Learning from Human Feedback (RLHF). Qwen2.5-Max is positioned by Alibaba as competitive with GPT-4o, Claude 3.5 Sonnet and DeepSeek V3 across general benchmarks, with reported leadership on Arena-Hard, LiveBench, LiveCodeBench and GPQA in the lab's own evaluations. The model is exclusively available via the Alibaba Cloud Model Studio API and through chat.qwenlm.ai; weights are not released. Function calling, JSON mode and vision (via separate Qwen2.5-VL-Max) are supported. The default context window is 32K tokens; longer contexts are available on request via Alibaba Cloud.

Parameters
Undisclosed (estimated several hundred billion total MoE parameters)
Context
32,768 tokens

Capabilities

  • Closed-weight MoE flagship of the Qwen2.5 family
  • Pretrained on 20T+ tokens, surpassing Qwen2.5-72B
  • Benchmarks competitive with GPT-4o, Claude 3.5 Sonnet, DeepSeek V3
  • Available exclusively via Alibaba Cloud Model Studio API
  • Function calling and JSON mode
  • Strong bilingual Chinese-English performance
  • Code generation across major programming languages
  • 32K default context window (longer on request)
  • Vision via separate Qwen2.5-VL-Max checkpoint
  • Cost-competitive with US frontier APIs
  • Best for: enterprise China-based deployments, bilingual chat, Alibaba Cloud customers.

Training & license

Pretrained on more than 20 trillion tokens of multilingual web text, code, books and scientific papers, with strong Chinese and English emphasis. Knowledge cutoff approximately late 2024. Post-training uses supervised fine-tuning and RLHF on curated instruction data.

License: Proprietary closed-weight commercial license via Alibaba Cloud Model Studio. Weights not released. Standard Alibaba Cloud commercial terms apply.

Safety testing: Alibaba applies standard SFT/RLHF safety alignment; the model filters Chinese politically sensitive topics in line with regulations. External third-party evaluations are limited.

Known limitations

  • Closed weights; no on-prem option
  • Filters Chinese political topics
  • Default 32K context shorter than Western flagships
  • API region availability concentrated in Asia
  • Limited third-party safety audits published
04

Pricing

Currently unavailable. There is no price for this model at the moment, so it cannot be run.

05

API

Call Qwen 2.5-Max with your Railwail API key. Use this model ID in the request:

No verified API example

The public API passes a different input format than this model needs. Use the playground above.

06

Specifications

Model ID
qwen-2-5-max
Developer
Other
Category
Text & chat
Input
Text
Output
Text
Context window
32,768 tokens
Max. output
8,192 tokens
Model size
Undisclosed (estimated several hundred billion total MoE parameters)
License
Proprietary closed-weight commercial license via Alibaba Cloud Model Studio. Weights not released. Standard Alibaba Cloud commercial terms apply.
Catalog entry updated
September 23, 2026

Tags

  • qwen
  • alibaba
  • moe
  • flagship
07

Use cases

What it is used for

  • Enterprise chat for Chinese market
  • Alibaba Cloud-hosted AI applications
  • Bilingual customer support
  • Document processing and summarisation
  • Coding assistants
  • RAG over enterprise corpora
08

Frequently asked questions

What is Qwen 2.5-Max?

Qwen 2.5-Max is a model by Other in the Text & chat category. It is listed on Railwail but cannot be run at the moment.

How much does Qwen 2.5-Max cost on Railwail?

Qwen 2.5-Max cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

What is the context window of Qwen 2.5-Max?

The context window of Qwen 2.5-Max holds 32,768 tokens. One response can be up to 8,192 tokens long.

How fast is Qwen 2.5-Max?

There are not enough measured runs of Qwen 2.5-Max on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

Is Qwen 2.5-Max better than Claude Fable 5.1?

That depends on the task. Qwen 2.5-Max (Other) and Claude Fable 5.1 (Anthropic) are both models in the Text & chat category. The comparison page shows their prices and specifications side by side.

Compare Qwen 2.5-Max and Claude Fable 5.1

Can I use Qwen 2.5-Max right now?

Currently unavailable. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.