Wire flash
Qwen releases Qwen-Audio-3.1 voice models, slashes prices up to 95% across entire line
Editorial responsibility
- No named human review is recorded for this page.
- Source reporting is collected, normalized, translated or condensed automatically when needed.
- Automatically published source-backed update
On September 23, Qwen (千问) officially released the Qwen-Audio-3.1 series of voice large models, according to a report from financial news outlet 财联社. The upgrade includes comprehensive improvements to three core models: automatic speech recognition (ASR), text-to-speech (TTS), and real-time voice interaction (Realtime). Additionally, Qwen launched two new models: the audio creation model Qwen-Audio-3.1-TTS-Next and the audio understanding model Qwen-Audio-3.1-ASR-Next. The five new models together form a complete audio capability stack covering 'understanding, generation, interaction, and creation.' To reduce user costs, Qwen has lowered prices across its entire Qwen-Audio voice model line. The price reductions include approximately 70% for TTS, about 85% for Realtime, and up to 95% for ASR.
Source report
September 23 (Caixin) — Qwen has officially launched the Qwen-Audio-3.1 series of voice large language models. This upgrade brings comprehensive improvements to three core models—Automatic Speech Recognition (ASR), Text-to-Speech (TTS), and Real-time Voice Interaction (Realtime)—and introduces two new models:
- Qwen-Audio-3.1-TTS-Next – an audio creation model
- Qwen-Audio-3.1-ASR-Next – an audio understanding model
The simultaneous release of five new voice models forms a complete audio capability stack covering "understanding, generation, interaction, and creation."
To further reduce user costs, Qwen-Audio has lowered prices across its entire voice model lineup:
- TTS: price reduced by approximately 70%
- Realtime: price reduced by approximately 85%
- ASR: price reduced by approximately 95%
Source
财联社Eastern
Part of this Story
Alibaba’s Qwen-Audio-3.1 Voice AI Models Launch with Price Cuts Up to 95%