Wire flash
Qwen releases Qwen-Audio-3.1 voice models, cuts prices across entire audio line
Editorial responsibility
- No named human review is recorded for this page.
- Source reporting is collected, normalized, translated or condensed automatically when needed.
- Automatically published source-backed update
On September 23, Jin10 Data reported that Qwen (千问) officially released the Qwen-Audio-3.1 series of voice large models. The upgrade includes comprehensive improvements to three core models: automatic speech recognition (ASR), text-to-speech (TTS), and real-time voice interaction (Realtime). Additionally, Qwen launched two new models: the audio creation model Qwen-Audio-3.1-TTS-Next and the audio understanding model Qwen-Audio-3.1-ASR-Next. The five new models form a complete audio capability stack covering 'understanding, generation, interaction, and creation.' To further reduce user costs, Qwen cut prices across its entire Qwen-Audio voice model line. Specifically, TTS prices were reduced by approximately 70%, Realtime by about 85%, and ASR by up to 95%.
Source report
On September 23, Qwen officially launched the Qwen-Audio-3.1 series of voice large models. This upgrade brings comprehensive improvements to three core models—Automatic Speech Recognition (ASR), Text-to-Speech (TTS), and Real-time Voice Interaction (Realtime)—and introduces two new models:
- Qwen-Audio-3.1-TTS-Next: An audio creation model
- Qwen-Audio-3.1-ASR-Next: An audio understanding model
The simultaneous release of five new voice models forms a complete audio capability stack covering "understanding, generation, interaction, and creation."
Price Reductions
To further lower user costs, Qwen has reduced prices across its entire Qwen-Audio voice model lineup:
- TTS: Approximately 70% reduction
- Realtime: Approximately 85% reduction
- ASR: Approximately 95% reduction
(Source: Qwen Large Model)
Source
金十数据Eastern
Part of this Story
Alibaba’s Qwen-Audio-3.1 Voice AI Models Launch with Price Cuts Up to 95%