Language modelOpen source
Jamba2-3B
AI21
- Released
- January 8, 2026
- Data date
- October 3, 2026
Jamba2-3B combines state-space and Transformer layers in a dense 3B model. AI21 targets answers grounded in supplied documents, including manuals and company policies. The compact version works without a reasoning model’s additional thinking process.
AI21 specifies 256K context tokens and releases the weights under Apache 2.0. Access is available through AI21 Studio and Hugging Face, with Jamba2-Mini joining the release as a larger MoE variant.
Measurements without matching peer values
These measurements have no matching peer values under the same test conditions. Their original values and sources remain available here.
IFBench (Show measurement, test conditions, and source)
- Source value
- 0.36
- Score
- 0.36
- Metric
- IFBench
- Unit
- ratio
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- AI21 announcement/HF chart
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. IFBench, Collie, IFEval and FACTS; Enterprise Reliability Score is their simple four-score average. Separate 100-prompt blind human side-by-side against Ministral3 14B, counterbalanced order.
- Context
- Source-specific observation; it is not a shared comparison cohort. AI21 self-reported comparison, not an external evaluator record
Collie (Show measurement, test conditions, and source)
- Source value
- 0.24
- Score
- 0.24
- Metric
- Collie
- Unit
- ratio
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- AI21 announcement/HF chart
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. IFBench, Collie, IFEval and FACTS; Enterprise Reliability Score is their simple four-score average. Separate 100-prompt blind human side-by-side against Ministral3 14B, counterbalanced order.
- Context
- Source-specific observation; it is not a shared comparison cohort. AI21 self-reported comparison, not an external evaluator record
IFEval (Show measurement, test conditions, and source)
- Source value
- 0.93
- Score
- 0.93
- Metric
- IFEval
- Unit
- ratio
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- AI21 announcement/HF chart
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. IFBench, Collie, IFEval and FACTS; Enterprise Reliability Score is their simple four-score average. Separate 100-prompt blind human side-by-side against Ministral3 14B, counterbalanced order.
- Context
- Source-specific observation; it is not a shared comparison cohort. AI21 self-reported comparison, not an external evaluator record
FACTS (Show measurement, test conditions, and source)
- Source value
- 0.54
- Score
- 0.54
- Metric
- FACTS
- Unit
- ratio
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- AI21 announcement/HF chart
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. IFBench, Collie, IFEval and FACTS; Enterprise Reliability Score is their simple four-score average. Separate 100-prompt blind human side-by-side against Ministral3 14B, counterbalanced order.
- Context
- Source-specific observation; it is not a shared comparison cohort. AI21 self-reported comparison, not an external evaluator record
Enterprise Reliability Score (average) (Show measurement, test conditions, and source)
- Source value
- 0.52
- Score
- 0.52
- Metric
- Enterprise Reliability Score (average)
- Unit
- ratio
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- AI21 announcement/HF chart
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. IFBench, Collie, IFEval and FACTS; Enterprise Reliability Score is their simple four-score average. Separate 100-prompt blind human side-by-side against Ministral3 14B, counterbalanced order.
- Context
- Source-specific observation; it is not a shared comparison cohort. AI21 self-reported comparison, not an external evaluator record
Sources and data date
Every statement links to its underlying documentation or leaderboard.
| Type | Evidence and data date |
|---|---|
| Research status | Research date October 4, 2026. 1 source URLs checked. This documents the inspected sources, not an exhaustive inventory of every publication. |
| Additional source | AI21 Jamba2 announcement (retrieved October 3, 2026; October 4, 2026) · AI21 Jamba2 announcement · Editorial description reviewed October 4, 2026 |
| Additional source | AI21 Jamba2-3B model card (retrieved October 3, 2026) · AI21 Jamba2-3B model card · Editorial description reviewed October 4, 2026 |