Garp Independent AI & technology journalism
Sunday, September 27, 2026 Sign In · Join Subscribe
Latest Don’t be fooled by this summer of AI hype 

AI news, research, models, robotics, chips, startups, and infrastructure coverage.

Updated daily

Home  /  AI News  /  Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent

AI News

Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent

Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio…

Alibaba’s AI team Qwen has released Qwen-Audio-3.1, a lineup of five models for speech recognition (ASR), text-to-speech (TTS), and real-time interaction. The ASR model improves multilingual and dialect recognition and automatically cleans up filler words and repetitions.

ASR-Next adds multi-speaker identification with timestamps and detects emotions, ambient sounds, and machine noise. TTS handles multilingual synthesis with natural cross-language voice transfer. Users control emotion, speed, and style through simple text prompts like “Read this with a sharp, commanding tone, demanding respect.” TTS-Next pairs a language model with a diffusion approach to generate voice, sound effects, and background audio in a single pass. The real-time model supports simultaneous speaking and listening with instant interruption. When it detects a low mood, it responds more slowly and with more empathy, according to Qwen. Alibaba is also slashing prices. TTS drops about 70 percent, Realtime roughly 85 percent, and ASR up to 95 percent. More details on the blog and on Qwen Cloud. Follow The Decoder for AI news, background stories and expert analyses.