Model index
Alibaba Cloud·Qwen

Qwen3.8-Omni-Flash-Realtime

The realtime interaction tier of Qwen3.8-Omni-Flash for camera and microphone scenarios, with text/audio output and multichannel audio.

CurrentLanguageAgentMultimodalWeights not released

QwenCloud released `qwen3.8-omni-flash-realtime` on 2026-09-21 through WebSocket, WebRTC, and AOQ for real-time audio/video interaction; official docs also list text/audio output, tool use, and up to 196,608 input tokens.

MODEL NUMBER103
First public date
2026-09-21
Organization
Alibaba Cloud
Family
Qwen
01

Specifications

Only published specifications are shown.

Context
196,608 tokens
Input modalities
Text · Image · Audio · Video
Output modalities
Text · Audio
Open weights
No
Status
Current
02

Overview

Overview

The realtime interaction tier of Qwen3.8-Omni-Flash for camera and microphone scenarios, with text/audio output and multichannel audio.

Why it mattered

It extends Qwen3.8 omni-modal capability from file understanding into real-time audio/video dialogue and MCP tool workflows.

03

Architecture & capabilities

Model capabilities

Native multimodalNative audioReal-time audioReal-time videoTool useVoice cloning

Connected tools

Web searchMCP tools
04

Release history

API launch

QwenCloud released qwen3.8-omni-flash-realtime for real-time audio/video interaction with text/audio output, multichannel audio, video aggregation, remote MCP tools, and WebSocket, WebRTC, and AOQ access.

View source
05

Related models

Select a card to view the related model.

06

Sources

Suggest a correction

Copy this template, add the field, proposed value, and primary source, and share it with the archive maintainer.