Skip to main content
Language modelActive

Mercury 2

Inception Labs

Released
February 24, 2026
Data date
October 3, 2026

Mercury 2 generates answers by refining text in parallel rather than through autoregressive token-by-token decoding. Inception Labs applies a diffusion approach, more commonly associated with image generation, to a reasoning language model.

The provider specifies a 128K-token context window. It is available through the Inception API and Mercury Chat, with Mercury 2.5 later expanding context to 260K tokens.

Specifications and access

SpecificationValue and source
Context window
128KSource
Architecture
Diffusion-based response generationSource
Capabilities
Tunable reasoning, native tool use, and schema-aligned JSONSource
Availability
API and Mercury ChatSource