Model index
Alibaba Cloud·Qwen

Qwen-Audio 3.1 Realtime Plus

The realtime speech tier of Qwen-Audio 3.1 for low-latency full-duplex dialogue, transcription, speech output, function calling, and web search.

CurrentLanguageMultimodalWeights not released

QwenCloud listed `qwen-audio-3.1-realtime-plus` on 2026-09-20 with full-duplex voice over WebSocket, AOQ, and WebRTC, retaining protocol compatibility with 3.0 Plus.

MODEL NUMBER104
First public date
2026-09-20
Organization
Alibaba Cloud
Family
Qwen
01

Specifications

Only published specifications are shown.

Context
262,144 tokens
Input modalities
Audio
Output modalities
Text · Audio
Open weights
No
Status
Current
02

Overview

Overview

The realtime speech tier of Qwen-Audio 3.1 for low-latency full-duplex dialogue, transcription, speech output, function calling, and web search.

Why it mattered

It advances Qwen’s audio line into a fuller voice-agent interface with additional system voices and a longer context.

03

Architecture & capabilities

Model capabilities

Native audioReal-time audioTool useVoice cloning

Connected tools

Web search

API features

Function calling
04

Release history

API launch

QwenCloud released qwen-audio-3.1-realtime-plus for full-duplex realtime voice with text/audio output, function calling, web search, and a 262,144-token context.

View source
05

Related models

Select a card to view the related model.

06

Sources

Suggest a correction

Copy this template, add the field, proposed value, and primary source, and share it with the archive maintainer.