Qwen 2.5 72B

Text & chatUnavailable
by Alibaba (Qwen)Model ID: qwen-2-5-72b

Alibaba's powerful open-source model. Excellent at coding, math, and multilingual tasks.

Status
Unavailable
Context
131,072 tokens
Max. output
4,096 tokens
Input โ†’ output
Text โ†’ Text
Developer
Alibaba (Qwen)
Updated
September 23, 2026

Qwen 2.5 72B is currently unavailable

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

Go to alternatives
01

Comparable models

All in this category
02

Playground

Try Qwen 2.5 72B

Chat

Currently unavailable

Currently unavailable.

The playground is disabled. You can find comparable models in the same category: Browse alternatives

Try Qwen 2.5 72B

Send a message. The answer arrives in full once the model is done (no streaming).

System prompt
Max. answer length (tokens)

This run

No price โ€“ currently unavailable.

New here?

10 free credits ($0.10) when you sign up with Google

Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits.

03

About Qwen 2.5 72B

TL;DRAs of September 23, 2026

Qwen 2.5 72B is a model by Alibaba (Qwen) in the Text & chat category. Qwen 2.5 72B is currently not available on Railwail. The context window holds 131,072 tokens, and one response can be up to 4,096 tokens long.

Background

About Alibaba Cloud (Qwen team)

Founded 2009 ยท Hangzhou, China

The Qwen (้€šไน‰ๅƒ้—ฎ, Tongyi Qianwen) team is the LLM research group inside Alibaba Cloud, the cloud-computing arm of Alibaba Group founded in 2009. Alibaba started large-language-model research at the Damo Academy and Tongyi Lab; the Qwen series became publicly available in 2023. Releases include Qwen-7B (Aug 2023, first Chinese open-weight foundation model from a major cloud), Qwen-14B/72B (late 2023), Qwen 1.5 (Feb 2024), Qwen2 family (Jun 2024), Qwen2.5 family (Sep 2024 with 0.5B/1.5B/3B/7B/14B/32B/72B sizes plus Qwen2.5-Coder, Qwen2.5-Math, Qwen2.5-VL), Qwen3 family in 2025 introducing thinking mode, and Qwen2.5-Max as the closed-weight flagship. The Qwen team is led by Junyang Lin and has published over a dozen widely cited technical reports. Models are released under the Tongyi Qianwen LICENSE (commercial-friendly for most use cases but with a >100M MAU clause requiring a separate license). Qwen has become the dominant open-weight model family for Chinese-language tasks and has the largest derivative-model ecosystem on HuggingFace by download volume.

Visit Alibaba Cloud (Qwen team)

Architecture

Decoder-only Transformer (dense, Grouped Query Attention)

Qwen2.5-72B was released by the Alibaba Qwen team in September 2024 as the flagship dense model of the Qwen2.5 family. It is a decoder-only Transformer with 72.7B parameters, 80 layers and Grouped Query Attention (GQA) with 64 query heads and 8 KV heads. The model was pretrained on a 18-trillion-token multilingual corpus emphasising Chinese, English and 27 other languages, plus heavy concentrations of math, code and long-form documents. Compared to Qwen2-72B (7T tokens), the 18T-token pretraining and improved data filtering pipeline meaningfully lifted knowledge, math and code performance. Qwen2.5-72B supports a 128K token (131,072) context window with YaRN scaling extension and ships in both base and Instruct variants. Post-training applied a multi-stage SFT pipeline on over 1M curated examples followed by Direct Preference Optimisation (DPO) and Group Relative Policy Optimisation (GRPO) for reasoning data. Qwen2.5-Coder-32B and Qwen2.5-Math-72B specialised siblings push code and math frontiers, and Qwen2.5-VL adds vision. Weights are released under the Tongyi Qianwen LICENSE (free commercial use unless >100M MAU). The Qwen ecosystem also publishes GGUF, AWQ, GPTQ quantised variants and is supported by vLLM, SGLang, llama.cpp, Ollama, MLX and Apple MLX.

Parameters
72.7B (dense)
Context
131,072 tokens

Capabilities

  • 72.7B dense parameters with Grouped Query Attention
  • Pretrained on 18T multilingual tokens
  • 128K context window with YaRN scaling
  • Strong multilingual performance: Chinese, English, Japanese, Korean and 27 more
  • Specialised siblings: Qwen2.5-Coder, Qwen2.5-Math, Qwen2.5-VL
  • Function calling and JSON mode supported
  • Open weights under Tongyi Qianwen LICENSE (commercial-friendly)
  • Massive ecosystem: vLLM, SGLang, llama.cpp, Ollama, MLX, HuggingFace
  • GGUF/AWQ/GPTQ quantised variants officially published
  • Strong math performance (Qwen2.5-Math variant beats GPT-4o on competition math)
  • Best for: open-weight Chinese/English chat, coding, on-prem enterprise, RAG.

Training & license

Pretrained on 18 trillion tokens of multilingual web text, code, books and scientific papers, with strong Chinese and English coverage. Post-training uses 1M+ curated SFT examples followed by DPO and GRPO. Data cutoff approximately mid-2024.

License: Tongyi Qianwen LICENSE: free commercial use for products with fewer than 100M monthly active users. >100M MAU requires a separate commercial license from Alibaba.

Safety testing: Qwen applies SFT-based safety alignment plus DPO with safety preferences. The model filters Chinese politically sensitive topics; third-party safety evaluations have flagged below-average refusal rates on Western dual-use prompts.

Known limitations

  • Filters Chinese political topics
  • Tongyi Qianwen LICENSE adds a >100M MAU clause
  • Vision requires the separate Qwen2.5-VL checkpoint
  • Knowledge cutoff mid-2024
  • Long context recall quality degrades beyond ~64K
04

Pricing

Currently unavailable. There is no price for this model at the moment, so it cannot be run.

05

API

Call Qwen 2.5 72B with your Railwail API key. Use this model ID in the request:

Currently unavailable

The model has no verified price or is deactivated; API calls are refused.

06

Specifications

Model ID
qwen-2-5-72b
Developer
Alibaba (Qwen)
Category
Text & chat
Input
Text
Output
Text
Context window
131,072 tokens
Max. output
4,096 tokens
Model size
72.7B (dense)
License
Tongyi Qianwen LICENSE: free commercial use for products with fewer than 100M monthly active users. >100M MAU requires a separate commercial license from Alibaba.
Catalog entry updated
September 23, 2026

Tags

  • open-source
  • coding
  • multilingual
07

Example prompts

Examples from the Railwail catalog. They were not generated live on this page.
  • Data Analysis

    Given monthly sales data [Jan: 12000, Feb: 15000, Mar: 14500, Apr: 18000, May: 22000, Jun: 19500], identify the trend and forecast July
    Show example answer

    The data shows a general upward trend with month-over-month growth averaging about 12.5%. There's a notable spike in May (+22.2%) followed by a correction in June (-11.4%), suggesting possible seasonality or a promotional event in May. Using linear regression on the six data points, the projected July sales are approximately 21,200. However, given the May-June pullback pattern, a more conservative estimate of 20,000-21,500 accounts for potential volatility.

  • Regex Help

    Write a regex pattern that validates email addresses and explain each part
    Show example answer

    Pattern: `^[a-zA-Z0-9._%+-]+@[a-zA-Z0-9.-]+\.[a-zA-Z]{2,}$` Breakdown: - `^` โ€“ Start of string - `[a-zA-Z0-9._%+-]+` โ€“ One or more valid local-part characters (letters, digits, dots, underscores, percent, plus, hyphen) - `@` โ€“ Literal @ symbol - `[a-zA-Z0-9.-]+` โ€“ Domain name (letters, digits, dots, hyphens) - `\.` โ€“ Literal dot before TLD - `[a-zA-Z]{2,}` โ€“ TLD must be at least 2 letters - `$` โ€“ End of string Note: This covers most common emails but won't match every valid RFC 5322 address.

08

Use cases

What it is used for

  • Chinese-English bilingual chat
  • Open-weight coding assistants
  • On-prem enterprise deployments
  • Multilingual document processing
  • Fine-tuning base for vertical models
  • Edge deployments via GGUF quantisation
09

Frequently asked questions

What is Qwen 2.5 72B?

Qwen 2.5 72B is a model by Alibaba (Qwen) in the Text & chat category. It is listed on Railwail but cannot be run at the moment.

How much does Qwen 2.5 72B cost on Railwail?

Qwen 2.5 72B cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

What is the context window of Qwen 2.5 72B?

The context window of Qwen 2.5 72B holds 131,072 tokens. One response can be up to 4,096 tokens long.

How fast is Qwen 2.5 72B?

There are not enough measured runs of Qwen 2.5 72B on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

Is Qwen 2.5 72B better than Claude Fable 5.1?

That depends on the task. Qwen 2.5 72B (Alibaba (Qwen)) and Claude Fable 5.1 (Anthropic) are both models in the Text & chat category. The comparison page shows their prices and specifications side by side.

Compare Qwen 2.5 72B and Claude Fable 5.1

Can I use Qwen 2.5 72B right now?

Currently unavailable. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.