AI21 Jamba 1.5 Large

Text & chatUnavailable
by AI21 LabsModel ID: jamba-1-5-large

AI21's flagship hybrid Mamba-Transformer model with a 256k context window for long-document tasks.

Status
Unavailable
Context
256,000 tokens
Max. output
4,096 tokens
Input โ†’ output
Text โ†’ Text
Developer
AI21 Labs
Updated
September 23, 2026

AI21 Jamba 1.5 Large is currently unavailable

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

Go to alternatives
01

Comparable models

All in this category
02

Playground

Try AI21 Jamba 1.5 Large

Input & output

Currently unavailable

Currently unavailable.

The playground is disabled. You can find comparable models in the same category: Browse alternatives

Try AI21 Jamba 1.5 Large
Output
The answer appears here.

This run

No price โ€“ currently unavailable.

New here?

10 free credits ($0.10) when you sign up with Google

Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits.

03

About AI21 Jamba 1.5 Large

TL;DRAs of September 23, 2026

AI21 Jamba 1.5 Large is a model by AI21 Labs in the Text & chat category. AI21 Jamba 1.5 Large is currently not available on Railwail. The context window holds 256,000 tokens, and one response can be up to 4,096 tokens long.

Background

About AI21 Labs

Founded 2017 ยท Tel Aviv, Israel

AI21 Labs is one of the earliest commercial LLM companies, founded in 2017 in Tel Aviv by Yoav Shoham (Stanford emeritus, AI pioneer), Ori Goshen and Amnon Shashua (Mobileye founder, ex-Intel SVP). AI21 built the Jurassic-1 (2021) and Jurassic-2 (2023) families and pioneered hybrid State-Space + Transformer architectures with Jamba in March 2024 โ€” the first production-scale Mamba-Transformer hybrid LLM. Jamba 1.5 followed in August 2024 at two scales: Mini (52B total / 12B active) and Large (398B total / 94B active). AI21 has raised over $336M from investors including Google, Nvidia, Walden Catalyst and Pitango, and serves enterprise customers through AI21 Studio, AWS Bedrock, Azure AI Studio, and Snowflake Cortex.

Visit AI21 Labs

Architecture

Hybrid Mamba-Transformer Mixture-of-Experts

Jamba 1.5 Large is a hybrid State-Space + Transformer Mixture-of-Experts model. The architecture interleaves Mamba (selective state-space) layers with standard self-attention layers in a 7:1 Mamba-to-Attention ratio across 72 blocks. Mixture-of-Experts is applied to MLP modules in attention blocks with 16 experts and top-2 routing, giving 94B active parameters out of 398B total. The Mamba layers handle long-range dependencies with O(N) memory while the attention layers preserve in-context retrieval quality, enabling a true 256,000-token effective context โ€” empirically validated on the RULER long-context benchmark, where pure-transformer 128K models degrade noticeably. The model uses a 64,000-token BPE tokeniser and supports nine languages (English, Spanish, French, Portuguese, Italian, Dutch, German, Arabic, Hebrew). Released August 2024 under the Jamba Open Model License with hosted access via AI21 Studio, AWS Bedrock, Azure AI Studio and Snowflake Cortex.

Parameters
398B total, 94B active per token (16 experts, top-2 routing)
Context
256,000 tokens

Capabilities

  • Hybrid Mamba+Transformer+MoE architecture
  • 398B total / 94B active parameters
  • 256K effective context โ€” best-in-class on RULER long-context benchmark
  • Constant memory per token from Mamba โ€” cheap long-context inference
  • Native function calling and JSON-mode structured output
  • Multilingual: English, Spanish, French, Portuguese, Italian, Dutch, German, Arabic, Hebrew
  • Open weights on Hugging Face under Jamba Open Model License
  • Best for: long-document analysis, many-document RAG, long-trace agents, cost-efficient enterprise long-context inference.

Training & license

Pretrained on trillions of tokens of web data, code, math, books and multilingual sources (exact figure not disclosed). Knowledge cutoff March 2024. Post-training is supervised fine-tuning plus preference optimisation; the SSAM (state-space attention mix) post-training adapts Mamba-state regularisation.

License: Jamba Open Model License. Permissive for research and commercial use with attribution and AUP compliance โ€” weaker than Apache 2.0 but more open than research-only licenses. Hosted commercial access via AI21 Studio, AWS Bedrock, Azure AI Studio and Snowflake Cortex.

Safety testing: AI21 publishes a model card with safety evaluations and limitations. Enterprise positioning emphasises grounded outputs and reduced hallucination over consumer-style chat safety tuning.

Known limitations

  • 398B total parameters need ~8x H100 for FP16 inference
  • No vision modality
  • Hybrid architecture has less community tooling โ€” some inference engines unsupported
  • Behind GPT-4o / Claude 3.5 Sonnet on hardest reasoning, code and math
  • Jamba Open Model License has acceptable-use restrictions and attribution requirements
04

Pricing

Currently unavailable. There is no price for this model at the moment, so it cannot be run.

05

API

Call AI21 Jamba 1.5 Large with your Railwail API key. Use this model ID in the request:

No verified API example

The public API passes a different input format than this model needs. Use the playground above.

06

Specifications

Model ID
jamba-1-5-large
Developer
AI21 Labs
Category
Text & chat
Input
Text
Output
Text
Context window
256,000 tokens
Max. output
4,096 tokens
Model size
398B total, 94B active per token (16 experts, top-2 routing)
License
Jamba Open Model License. Permissive for research and commercial use with attribution and AUP compliance โ€” weaker than Apache 2.0 but more open than research-only licenses. Hosted commercial access via AI21 Studio, AWS Bedrock, Azure AI Studio and Snowflake Cortex.
Catalog entry updated
September 23, 2026

Tags

  • ai21
  • long-context
  • mamba
  • hybrid
  • open-weights
07

Use cases

What it is used for

  • Whole-book and long-contract analysis
  • Many-document RAG with large retrieved windows
  • Long-trace agentic workflows
  • Cost-efficient enterprise long-context inference
  • Multilingual document understanding (9 languages)
  • Research on SSM-Transformer hybrid architectures
08

Frequently asked questions

What is AI21 Jamba 1.5 Large?

AI21 Jamba 1.5 Large is a model by AI21 Labs in the Text & chat category. It is listed on Railwail but cannot be run at the moment.

How much does AI21 Jamba 1.5 Large cost on Railwail?

AI21 Jamba 1.5 Large cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

What is the context window of AI21 Jamba 1.5 Large?

The context window of AI21 Jamba 1.5 Large holds 256,000 tokens. One response can be up to 4,096 tokens long.

How fast is AI21 Jamba 1.5 Large?

There are not enough measured runs of AI21 Jamba 1.5 Large on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

Is AI21 Jamba 1.5 Large better than Claude Fable 5.1?

That depends on the task. AI21 Jamba 1.5 Large (AI21 Labs) and Claude Fable 5.1 (Anthropic) are both models in the Text & chat category. The comparison page shows their prices and specifications side by side.

Compare AI21 Jamba 1.5 Large and Claude Fable 5.1

Can I use AI21 Jamba 1.5 Large right now?

Currently unavailable. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.