chiprook
← AI
AISeptember 29, 2026, 11:58

Alibaba Qwen Releases Qwen-Audio-3.1-Realtime for Voice Agents

Qwen launched the Qwen-Audio-3.1 stack of five models spanning ASR, TTS and realtime interaction. The flagship Qwen-Audio-3.1-Realtime is a full-duplex speech model with function calling and web search, a 262K-token context and $6.4 per 1M audio input tokens. No open weights; access is via the QwenCloud API.

Alibaba Qwen Releases Qwen-Audio-3.1-Realtime for Voice Agents
#Alibaba#Qwen
Read next
AI

Grok Voice Realtime: xAI’s Audio-to-Audio Model Explained

AI

Alibaba releases open-weights AI agent claiming frontier co-work scores at 3B active parameters

AI

Tencent Releases AI Image Model to Catch ByteDance, Alibaba

AI

Alibaba Releases Qwen-Image-2.1, 7B Open-Weight Image Model