AI21 Labs Pricing
AI21 Labs charges between $0.400 and $8.00 per 1 million output tokens, depending on the model. Jamba 1.5 is the cheapest at $0.400/1M output and $0.200/1M input; Jamba 1.5 Large is the most expensive at $8.00/1M output. The largest context window is 256K tokens (Jamba 1.5). Prices are USD, current as of July 20, 2026.
Last updated · synced weekly from the upstream model catalog
Cheapest input
$0.200/1M
Jamba 1.5
Cheapest output
$0.400/1M
Jamba 1.5
Longest context
256K
Jamba 1.5
Models priced
9
Avg $3.78/1M output
Pricing by model
All prices in USD per 1 million tokens. All 9 priced models, cheapest first.
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
| Jamba 1.5 | $0.200 | $0.400 | 256K |
| Jamba 1.5 Mini | $0.200 | $0.400 | 256K |
| Jamba 1.5 Mini 001 | $0.200 | $0.400 | 256K |
| Jamba Mini 1.6 | $0.200 | $0.400 | 256K |
| Jamba Mini 1.7 | $0.200 | $0.400 | 256K |
| Jamba 1.5 Large | $2.00 | $8.00 | 256K |
| Jamba 1.5 Large 001 | $2.00 | $8.00 | 256K |
| Jamba Large 1.6 | $2.00 | $8.00 | 256K |
| Jamba Large 1.7 | $2.00 | $8.00 | 256K |
Frequently asked questions
How much does AI21 Labs cost per 1M tokens?
AI21 Labs pricing ranges from $0.400 to $8.00 per 1 million output tokens across 9 models. Input tokens are cheaper than output on every model. Rates current as of July 20, 2026.
What is the cheapest AI21 Labs model?
Jamba 1.5 is the cheapest AI21 Labs model on both axes — $0.200 per 1M input tokens and $0.400 per 1M output. It is the right default for high-volume, low-complexity work such as classification, extraction and routing.
What is the largest AI21 Labs context window?
Jamba 1.5 has the largest context window in the AI21 Labs catalog at 256K input tokens, with up to 256K output tokens per response.
How do I calculate my actual AI21 Labs bill?
Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.
Other providers
What will this actually cost you?
Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.
Open the free cost calculator