MiniMax-01

Text & chatUnavailable
by MiniMaxModel ID: minimax-01

MiniMax's 456B hybrid lightning-attention model with native 4M-token context. Industry-leading long-context.

Status
Unavailable
Context
4,096,000 tokens
Max. output
16,384 tokens
Input โ†’ output
Text โ†’ Text
Developer
MiniMax
Updated
23 September 2026

MiniMax-01 is currently unavailable

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

Go to alternatives
01

Comparable models

All in this category
02

Playground

Try MiniMax-01

Input & output

Currently unavailable

Currently unavailable.

The playground is disabled. You can find comparable models in the same category: Browse alternatives

Try MiniMax-01
Output
The answer appears here.

This run

No price โ€“ currently unavailable.

New here?

10 free credits (US$0.10) when you sign up with Google

Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits.

03

About MiniMax-01

TL;DRAs of 23 September 2026

MiniMax-01 is a model by MiniMax in the Text & chat category. MiniMax-01 is currently not available on Railwail. The context window holds 4,096,000 tokens, and one response can be up to 16,384 tokens long.

Background

About MiniMax

Founded 2021 ยท Shanghai, China

MiniMax (ไธŠๆตท็จ€ๅฎ‡็ง‘ๆŠ€, Shanghai Xiyu Technology) is a Shanghai-based AI startup founded in late 2021 by Yan Junjie (former vice president of SenseTime). The company has built a portfolio of consumer AI products including the Talkie and Hailuo AI chatbots, abab text models, and the Hailuo Video generative video line. MiniMax has raised over $850M from investors including Alibaba, Tencent, Hillhouse Capital and IDG Capital, with a 2024 valuation above $2.5B. MiniMax-01, released January 2025, was the company's first open-weight frontier release and introduced lightning attention at scale โ€” a linear-attention variant โ€” with a 4M token training context and 1M context inference window.

Visit MiniMax

Architecture

Hybrid Lightning-Attention + Softmax-Attention Mixture-of-Experts

MiniMax-01 (MiniMax-Text-01 base plus MiniMax-VL-01 vision variant) combines lightning attention โ€” MiniMax's linear-attention design โ€” with periodic softmax attention layers (every 8th layer is full softmax). The architecture has 80 layers with 6,144 hidden size, and MoE feed-forwards with 32 experts and top-2 routing yielding 45.9B active out of 456B total. The hybrid design enables 4M-token effective training context and 1M-token inference context, with near-linear compute per token in the lightning-attention layers. MiniMax describes MiniMax-01 in its technical paper as the first production-scale linear-attention LLM. The VL variant adds 336x336 image patch encoding for vision input. Released January 2025 under the MiniMax Model License with open weights on Hugging Face alongside hosted API access via the MiniMax Open Platform.

Parameters
456B total, 45.9B active per token (32 experts, top-2 routing)
Context
4,000,000 tokens

Capabilities

  • Industry-leading 1M-4M token context window
  • First production-scale lightning (linear) attention LLM
  • 456B total / 45.9B active parameters
  • Strong needle-in-haystack performance reported in the technical paper
  • Competitive with GPT-4o on standard benchmarks (MMLU, GSM8K, HumanEval)
  • Open weights released โ€” first frontier-scale lightning-attention model in the open
  • Vision variant (MiniMax-VL-01) supports image inputs
  • Multilingual with strong Chinese and English performance
  • Best for: ultra-long-context analysis, Chinese-language applications, long-form agentic workflows, research on linear attention.

Training & license

Pretrained on trillions of tokens of multilingual web data with heavy Chinese and English representation, code and math; the VL variant adds image-text pairs. Exact data composition is partially described in the technical paper. Knowledge cutoff approximately late 2024.

License: MiniMax Model License. Permits commercial use with acceptable-use restrictions; products at scale may require registration with MiniMax. Review the license file on Hugging Face before deployment.

Safety testing: MiniMax conducted internal safety evaluation. The model filters politically sensitive topics consistent with Chinese regulations. No public third-party red-team report.

Known limitations

  • 456B total parameters โ€” substantial GPU memory required for self-hosting
  • Lightning attention has less mature kernel support than softmax
  • Behind o1 / R1 / Claude 4 on hardest reasoning tasks
  • Vision quality below GPT-4o and Claude 3.5 Sonnet
  • Filters politically sensitive topics consistent with Chinese regulations
04

Pricing

Currently unavailable. There is no price for this model at the moment, so it cannot be run.

05

API

Call MiniMax-01 with your Railwail API key. Use this model ID in the request:

No verified API example

The public API passes a different input format than this model needs. Use the playground above.

06

Specifications

Model ID
minimax-01
Developer
MiniMax
Category
Text & chat
Input
Text
Output
Text
Context window
4,096,000 tokens
Max. output
16,384 tokens
Model size
456B total, 45.9B active per token (32 experts, top-2 routing)
License
MiniMax Model License. Permits commercial use with acceptable-use restrictions; products at scale may require registration with MiniMax. Review the license file on Hugging Face before deployment.
Catalog entry updated
23 September 2026

Tags

  • minimax
  • long-context
  • lightning-attention
  • open-weights
  • 4m-context
07

Use cases

What it is used for

  • Ultra-long-context document and codebase analysis
  • Long-trace agentic workflows
  • Chinese-language consumer and enterprise chat
  • Long-form video transcript reasoning
  • Research on linear-attention LLMs
  • Multi-book or multi-PDF synthesis
08

Frequently asked questions

What is MiniMax-01?

MiniMax-01 is a model by MiniMax in the Text & chat category. It is listed on Railwail but cannot be run at the moment.

How much does MiniMax-01 cost on Railwail?

MiniMax-01 cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

What is the context window of MiniMax-01?

The context window of MiniMax-01 holds 4,096,000 tokens. One response can be up to 16,384 tokens long.

How fast is MiniMax-01?

There are not enough measured runs of MiniMax-01 on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

Is MiniMax-01 better than Claude Fable 5.1?

That depends on the task. MiniMax-01 (MiniMax) and Claude Fable 5.1 (Anthropic) are both models in the Text & chat category. The comparison page shows their prices and specifications side by side.

Compare MiniMax-01 and Claude Fable 5.1

Can I use MiniMax-01 right now?

Currently unavailable. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = US$0.01.