Skip to main content
Audio modelActive

Qwen3.8-Omni-Flash-Realtime

Alibaba

Released
September 21, 2026
Data date
October 3, 2026

Qwen3.8-Omni-Flash-Realtime connects live text, audio, and video input to text and audio output. Its realtime documentation lists WebSocket, WebRTC, and AOQ transports; function calling and MCP form the documented action path.

The variant has a 196,608-token input context and up to 65,536 output tokens. Rates differ between Beijing and Singapore. Speech output is billed for both audio and its corresponding text. No separate new weights checkpoint is published.

Specifications and access

SpecificationValue and source
Model class
Omnimodal real-time modelSource
Access
Alibaba Cloud Model Studio APISource
Input
Text, audio, and videoSource
Output
Text and audioSource
API model ID
qwen3.8-omni-flash-realtimeSource
Audio history
Up to 100 audio turns or 600 seconds; video up to 50 turns or 240 secondsSource
Transport
WebSocket, WebRTC, or AOQSource
Classification
Real-time API variant without separately published new weightsSource
Singapore rate limit
60 requests and 2M tokens per minuteSource