Model catalog

AI21 Labs Models

Browse the AI models associated with AI21 Labs. Compare current and historical models by family, capabilities, context window, availability and intended use.

10 models tracked
10 Total models
6 Model families
3 Model types
10 Current / accessible
All models

AI21 Labs model catalog

Long-context text generation, open-weight research, experimentation, and private or self-hosted deployment

Type General Purpose
Context 256K
Reasoning 6/10
Speed 8/10
Status

Legacy open-weight model; downloadable and accessible through AI21 Labs' official Hugging Face repository

View model →

Long-document analysis, grounded generation, enterprise RAG, document summarization, information extraction, private deployment, and multilingual text workflows.

Type General Purpose
Context 256K
Reasoning 7/10
Speed 8/10
Fine-tuning
Status

Current and available; the API endpoint jambalarge-1.7 points to the dated snapshot jambalarge-1.7-2025-07.

Input $2 per 1M input tokens
Output $8 per 1M output tokens
View model →

Long-context document analysis, RAG, structured generation, function calling, multilingual enterprise assistants, and private deployment

Type General Purpose
Context 256K
Reasoning 5/10
Speed 8/10
Tool use Fine-tuning Streaming
Status

Legacy; superseded by newer Jamba releases, including AI21-Jamba-Mini-1.7. Public model weights remain available.

View model →

Long-context document analysis, retrieval-augmented generation, enterprise assistants, structured text generation, multilingual workflows, and self-hosted deployments

Type General Purpose
Context 262K
Reasoning 7/10
Speed 8/10
Tool use Fine-tuning Streaming
Status

Legacy and superseded by newer Jamba Large releases; downloadable open-weight checkpoint remains available

View model →

Long-context retrieval-augmented generation, enterprise document analysis, grounded question answering, structured extraction, classification, and private deployment

Type General Purpose
Context 256K
Reasoning 6/10
Speed 7/10
Tool use Fine-tuning Streaming
Status

Available; older Jamba 1.6 generation with Jamba 1.7 available as a newer successor

Input $2 per 1 million input tokens
Output $8 per 1 million output tokens
View model →

Long-context RAG, grounded question answering, enterprise document processing, classification, structured text generation, function calling and privacy-sensitive private deployments.

Type General Purpose
Context 256K
Reasoning 6/10
Speed 8/10
Tool use Fine-tuning Streaming
Status

Legacy/open-weight model; downloadable from AI21's official Hugging Face repository and available for self-managed deployment, but not featured among AI21's current primary Jamba models as of September 25, 2026.

View model →

Long-document analysis, enterprise RAG, grounded question answering, structured text generation, private deployment, and cost-sensitive text workflows

Type General Purpose
Context 256K
Reasoning 5/10
Speed 7/10
Tool use Fine-tuning Streaming
Status

Available as an AI21 open-weight model; hosted API availability is not currently verified

Input $0.20 per 1 million input tokens when offered through AI21-hosted inference
Output $0.40 per 1 million output tokens when offered through AI21-hosted inference
View model →

Long-context RAG, grounded enterprise question answering, document extraction, local inference, on-device assistants, and lightweight agent workflows

Type Lightweight
Context 262K
Reasoning 5/10
Speed 8/10
Tool use Streaming
Status

Current and available; open-weight release

View model →

Long-context enterprise question answering, grounded generation, document analysis, instruction-heavy workflows, RAG systems and self-hosted deployments

Type General Purpose
Context 256K
Reasoning 6/10
Speed 8/10
Streaming
Status

Current open-weight model; available through Hugging Face and AI21 Studio

View model →

Local reasoning, long-context document analysis, private RAG, extraction, coding assistance, and lightweight agent controllers

Type Reasoning
Context 256K
Reasoning 7/10
Speed 8/10
Tool use Streaming
Status

Current open-weight model; available for download and local inference

Input No official AI21 hosted API price; self-hosted/open-weight model
Output No official AI21 hosted API price; self-hosted/open-weight model
View model →