25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / AI21 Labs / Jamba / Jamba Large

Jamba Large

AI21's hybrid SSM-Transformer long-context model for enterprise agents and grounded generation

Line

Jamba

Weights

Open

Released

2025-07-01

Context

256K

Input

$2.00 / 1M

Output

$8.00 / 1M

Max output

4K

Coverage

not covered yet

Overview

Jamba Large is AI21 Labs' hybrid SSM-Transformer long-context model, built for enterprise agents and grounded generation. It is the first version of the jamba line in the registry, released on 2025-07-01. It has a context window of 256000 tokens and a maximum output of 4096 tokens. Pricing starts at $2.00 per million input tokens and $8.00 per million output tokens. The weights are open. It supports tool calling, accepts and returns text only, and does not have a reasoning mode. Its knowledge cutoff is 2024-08-22.

Specs

Released2025-07-01
LineAI21 Labs · Jamba
WeightsOpen
Context256K tokens
Max output4K tokens
Inputtext
Outputtext
ReasoningNo
Tool callingYes
Knowledge cutoff2024-08-22
API idjamba-large

Pricing

Input$2.00 / 1M tokens
Cached inputnot stated
Output$8.00 / 1M tokens

First version of its line in the registry.

Strengths

  • 256K token context window for long documents
  • Hybrid SSM-Transformer architecture for efficiency
  • Tool calling for enterprise agent workflows
  • Open weights for self-hosting and customization
  • Grounded generation for factual responses

Best for

  • Reach for it for long-context enterprise document analysis
  • Reach for it for building agents that call external tools
  • Reach for it for grounded generation over large corpora
  • Reach for it for text-only tasks with high token volumes

How to access

ProviderModel id
AI21 Labsjamba-large

1 host serve this model at the price above · the maker's documentation

Jamba: every version

VersionReleasedContextInput
Jamba MiniCurrent2026-01-01256K$0.20
Jamba Large2025-07-01256K$2.00

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window and output limit?
The context window is 256000 tokens and the maximum output is 4096 tokens per request.
Is the model open weights and what does it cost?
Yes, the weights are open. Pricing starts at $2.00 per million input tokens and $8.00 per million output tokens.
Does it support tool calling or reasoning?
It supports tool calling, but it does not have a reasoning mode. It accepts and returns text only.
Open weights256K context