Skip to main content
Language modelOpen source

Jamba Reasoning 3B

AI21

Released
October 8, 2025
Data date
October 3, 2026

Jamba Reasoning 3B pairs reasoning with a memory-efficient SSM-Transformer architecture. AI21 reports a KV cache eight times smaller than that of a conventional Transformer. At a 32K-token context, the provider measures 40 output tokens per second on an M3 MacBook Pro.

AI21 releases the 3B model under Apache 2.0 with a 256K context window. The provider documents local use through LM Studio and llama.cpp, among other options.

Specifications and access

SpecificationValue and source
Release
October 8, 2025Source
Context window
256K, with processing up to 1M tokens describedSource
License
Apache 2.0Source