Skip to main content
Language modelEnterprise access

Mercury Voice

Inception

Released
September 29, 2026
Data date
October 3, 2026

Mercury Voice targets the time to first answer token in voice agents. In its own tests on customer-service prompts, Inception reports a median of 320 milliseconds after reasoning. Despite its name, this diffusion LLM generates text; other components handle speech recognition and audio output.

Inception offers the checkpoint through its own API for enterprise customers. It handles 128K context tokens and up to 50K output tokens.

Specifications and access

SpecificationValue and source
Model class
Diffusion LLM for voice agentsSource
Context window
128K tokens according to InceptionSource
Maximum output
50K tokens according to InceptionSource
Output
Text, not text-to-speech audioSource
Access
Inception API for enterprise customersSource