Skip to main content
ModelsBusinessTools

Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent

Source: The Decoder (opens in a new tab) · Matthias Bastian

Intel Summary

Alibaba's Qwen AI team has launched Qwen-Audio-3.1, a suite of five models designed for automatic speech recognition (ASR), text-to-speech (TTS), and real-time interaction. The updated lineup introduces improved multilingual and dialect recognition, automated filler word removal, timestamped multi-speaker identification, and acoustic emotion and noise detection, alongside price cuts of up to 95 percent.

Why It Matters

Steep price reductions combined with advanced multi-speaker diarization and acoustic environment detection lower operational hurdles for voice agent and transcription deployments. The move intensifies commercial pricing pressure on speech-to-text and voice generation API providers across the AI ecosystem.

Part of an ongoing development

Source

Alibaba launches Qwen-Audio-3.1 and reduces AI audio pricing

Alibaba's Qwen AI team has launched Qwen-Audio-3.1, a suite of five models designed for automatic speech recognition (ASR), text-to-speech (TTS), and real-time interaction. The updated lineup introduces improved multilingual and dialect recognition, automated filler word removal, timestamped multi-speaker identification, and acoustic emotion and noise detection, alongside price cuts of up to 95 percent. Claims are as reported; this summary makes no determination about accuracy or significance.

Organizations & Entities