Skip to main content
Audio modelActive

MiniMax Speech 2.8

MiniMax

Released
January 23, 2026
Data date
October 3, 2026
Model class

MiniMax Speech 2.8 controls expressive speech synthesis through native sound tags and voice cloning from a ten-second reference. The family produces multilingual output and explicitly names Mandarin and Japanese among its supported languages.

The API offers the identifiers speech-2.8-hd and speech-2.8-turbo. MiniMax lists $100 per 1M characters for HD and $60 for Turbo for both synchronous and asynchronous speech synthesis. Voice design and cloning are billed separately.

Specifications and access

SpecificationValue and source
Model class
Speech synthesisSource
Access
MiniMax APISource
Input
Text and a short reference recordingSource
Output
AudioSource
API model ID
speech-2.8-hd, speech-2.8-turboSource
Voice cloning
High-fidelity cloning from a ten-second referenceSource
Languages
Mandarin, Japanese, and additional multilingual speech optionsSource
Sound tags
Native sound tags for expressive outputSource
Family
Family profile with distinct HD and Turbo API variantsSource