AI21 Jamba 1.5 Mini

Text & chatUnavailable
by AI21 LabsModel ID: jamba-1-5-mini

Cost-efficient hybrid Mamba-Transformer model with 256k context. Tuned for high-throughput RAG.

Status
Unavailable
Context
256,000 tokens
Max. output
4,096 tokens
Input β†’ output
Text β†’ Text
Developer
AI21 Labs
Updated
September 23, 2026

AI21 Jamba 1.5 Mini is currently unavailable

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

Go to alternatives
01

Comparable models

All in this category
02

Playground

Try AI21 Jamba 1.5 Mini

Input & output

Currently unavailable

Currently unavailable.

The playground is disabled. You can find comparable models in the same category: Browse alternatives

Try AI21 Jamba 1.5 Mini
Output
The answer appears here.

This run

No price – currently unavailable.

New here?

10 free credits ($0.10) when you sign up with Google

Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits.

03

About AI21 Jamba 1.5 Mini

TL;DRAs of September 23, 2026

AI21 Jamba 1.5 Mini is a model by AI21 Labs in the Text & chat category. AI21 Jamba 1.5 Mini is currently not available on Railwail. The context window holds 256,000 tokens, and one response can be up to 4,096 tokens long.

Background

About AI21 Labs

Founded 2017 Β· Tel Aviv, Israel

AI21 Labs is an Israeli LLM pioneer founded in 2017 by Yoav Shoham (Stanford emeritus), Ori Goshen and Amnon Shashua (Mobileye founder). Long active in commercial LLMs (Jurassic-1, Jurassic-2), AI21 pioneered the hybrid State-Space + Transformer Jamba architecture in March 2024. Jamba 1.5 Mini is the small-and-fast variant of the August 2024 Jamba 1.5 release, optimised for long-context inference at production cost points and tractable on a single H100 80GB. AI21 has raised over $336M from investors including Google, Nvidia, Walden Catalyst and Pitango.

Visit AI21 Labs

Architecture

Hybrid Mamba-Transformer Mixture-of-Experts

Jamba 1.5 Mini uses the same hybrid SSM+Transformer+MoE recipe as Jamba 1.5 Large at smaller scale. Across 32 blocks each block alternates Mamba (selective state-space) and self-attention layers in a 7:1 ratio. MLPs are MoE with 16 experts and top-2 routing, giving 12B active parameters out of 52B total. The 12B active count means inference cost is competitive with dense 12B models, while Mamba layers provide constant-memory long-context scaling. The model uses a 64,000-token BPE tokeniser and supports the same nine languages as Jamba 1.5 Large (English, Spanish, French, Portuguese, Italian, Dutch, German, Arabic, Hebrew). Released August 2024 under the Jamba Open Model License with hosted access via AI21 Studio, AWS Bedrock, Azure AI Studio and Snowflake Cortex.

Parameters
52B total, 12B active per token (16 experts, top-2 routing)
Context
256,000 tokens

Capabilities

  • Hybrid Mamba+Transformer+MoE architecture at compact scale
  • 52B total / 12B active parameters
  • 256K effective context
  • Fits on a single H100 80GB at FP16 or A100 at INT8
  • Native function calling and JSON-mode output
  • Multilingual (9 languages)
  • Open weights under Jamba Open Model License
  • Best for: cheap long-context summarisation, RAG with large retrieval windows, single-GPU enterprise pilots.

Training & license

Same data mixture and methodology as Jamba 1.5 Large: trillions of tokens of web, code, math, books and multilingual sources, knowledge cutoff March 2024, followed by supervised fine-tuning and preference optimisation.

License: Jamba Open Model License. Permits research and commercial use with attribution and AUP compliance. Hosted access via AI21 Studio, AWS Bedrock, Azure AI Studio and Snowflake Cortex.

Safety testing: AI21 publishes a model card with safety and limitation disclosures. Positioning is enterprise-focused (grounded, structured outputs) rather than consumer-chat safety tuning.

Known limitations

  • Lower quality than Jamba 1.5 Large or Mixtral 8x7B Instruct on hard reasoning
  • No vision modality
  • Limited community ecosystem β€” fewer inference engines support hybrid Mamba+attention
  • Behind dense 70B-class instructs on benchmark depth
  • Multilingual coverage narrower than Command R, Aya or Mistral Large
04

Pricing

Currently unavailable. There is no price for this model at the moment, so it cannot be run.

05

API

Call AI21 Jamba 1.5 Mini with your Railwail API key. Use this model ID in the request:

No verified API example

The public API passes a different input format than this model needs. Use the playground above.

06

Specifications

Model ID
jamba-1-5-mini
Developer
AI21 Labs
Category
Text & chat
Input
Text
Output
Text
Context window
256,000 tokens
Max. output
4,096 tokens
Model size
52B total, 12B active per token (16 experts, top-2 routing)
License
Jamba Open Model License. Permits research and commercial use with attribution and AUP compliance. Hosted access via AI21 Studio, AWS Bedrock, Azure AI Studio and Snowflake Cortex.
Catalog entry updated
September 23, 2026

Tags

  • ai21
  • long-context
  • mamba
  • hybrid
  • open-weights
  • cost-efficient
07

Use cases

What it is used for

  • Cheap long-context summarisation
  • RAG with large retrieved windows
  • Single-GPU enterprise pilots
  • Function-calling agents
  • Multilingual document processing
  • Cost-efficient AI21 hosted inference
08

Frequently asked questions

What is AI21 Jamba 1.5 Mini?

AI21 Jamba 1.5 Mini is a model by AI21 Labs in the Text & chat category. It is listed on Railwail but cannot be run at the moment.

How much does AI21 Jamba 1.5 Mini cost on Railwail?

AI21 Jamba 1.5 Mini cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

What is the context window of AI21 Jamba 1.5 Mini?

The context window of AI21 Jamba 1.5 Mini holds 256,000 tokens. One response can be up to 4,096 tokens long.

How fast is AI21 Jamba 1.5 Mini?

There are not enough measured runs of AI21 Jamba 1.5 Mini on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

Is AI21 Jamba 1.5 Mini better than Claude Fable 5.1?

That depends on the task. AI21 Jamba 1.5 Mini (AI21 Labs) and Claude Fable 5.1 (Anthropic) are both models in the Text & chat category. The comparison page shows their prices and specifications side by side.

Compare AI21 Jamba 1.5 Mini and Claude Fable 5.1

Can I use AI21 Jamba 1.5 Mini right now?

Currently unavailable. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.