Overview
The realtime speech tier of Qwen-Audio 3.1 for low-latency full-duplex dialogue, transcription, speech output, function calling, and web search.
The realtime speech tier of Qwen-Audio 3.1 for low-latency full-duplex dialogue, transcription, speech output, function calling, and web search.
QwenCloud listed `qwen-audio-3.1-realtime-plus` on 2026-09-20 with full-duplex voice over WebSocket, AOQ, and WebRTC, retaining protocol compatibility with 3.0 Plus.
Only published specifications are shown.
The realtime speech tier of Qwen-Audio 3.1 for low-latency full-duplex dialogue, transcription, speech output, function calling, and web search.
It advances Qwen’s audio line into a fuller voice-agent interface with additional system voices and a longer context.
QwenCloud released qwen-audio-3.1-realtime-plus for full-duplex realtime voice with text/audio output, function calling, web search, and a 262,144-token context.
View sourceSelect a card to view the related model.
Copy this template, add the field, proposed value, and primary source, and share it with the archive maintainer.