Skip to main content
Model ReleaseNew

Alibaba launches Qwen-Audio-3.1 and reduces AI audio pricing

Alibaba's Qwen AI team has launched Qwen-Audio-3.1, a suite of five models designed for automatic speech recognition (ASR), text-to-speech (TTS), and real-time interaction. The updated lineup introduces improved multilingual and dialect recognition, automated filler word removal, timestamped multi-speaker identification, and acoustic emotion and noise detection, alongside price cuts of up to 95 percent. Claims are as reported; this summary makes no determination about accuracy or significance.

First detected
Sep 23, 2026
Last updated
Sep 23, 2026

Newly detected

This development was detected recently and reporting may still arrive.

Follow this development to see meaningful updates as new evidence emerges.

Save keeps this for later. Follow tracks meaningful changes as new evidence emerges — it shapes your Following Feed, alerts, and digest eligibility, and doesn't promise an instant notification.

Why it matters

Steep price reductions combined with advanced multi-speaker diarization and acoustic environment detection lower operational hurdles for voice agent and transcription deployments. The move intensifies commercial pricing pressure on speech-to-text and voice generation API providers across the AI ecosystem.

Coverage

Primary/vendor sources vs independent reporting

Primary / vendor source: information published directly by the company, organization, government body or project involved. Useful as a primary source, but not independent confirmation.

Independent reporting: reporting or analysis from a source independent of the organization making the underlying claim.

How this developed

  1. Sep 23, 2026

    1. Development detected

    2. New reporting added