Qwen releases Qwen-Audio-3.1, with significant price cuts across its entire voice model lineup
BIBIBI
AT A GLANCE
Alibaba’s Qwen released five Qwen-Audio-3.1 voice models and reduced prices for all voice models.
Article
Qwen releases Qwen-Audio-3.1, reducing Qwen-Audio prices across its full range of speech models
PANews September 23 news, Alibaba Tongyi Qwen officially released Qwen-Audio-3.1 series speech foundation models. This upgrade not only comprehensively advances the core models of speech recognition, speech synthesis, and real-time speech interaction, but also introduces the new audio creation model Qwen-Audio-3.1-TTS-Next and audio understanding model Qwen-Audio-3.1-ASR-Next. new speech models were launched simultaneously, forming a complete audio capability stack covering “understanding-generation-interaction-creation.” To further reduce user costs,Qwen-Audio has reduced prices across its full range of speech models, among which TTS was reduced by approximately 70%,Realtime was reduced by approximately 85%,ASR was reduced by as much as 95%。
Alibaba Tongyi Qwen officially releases the Qwen-Audio-3.1 series of large voice models.
02
This release includes new voice models.
03
The new models include Qwen-Audio-3.1-TTS-Next and Qwen-Audio-3.1-ASR-Next。
04
Related capabilities cover automatic speech recognition, speech synthesis, real-time voice interaction, audio creation, and audio understanding.
05
Qwen-Audio Prices for all voice models have been reduced, among which TTS saw a price reduction of approximately 70%,Realtime saw a reduction of approximately 85%,ASR saw a reduction of up to 95%。
AI-assisted interpretation
The following is analysis, separate from reported facts. Verify important claims independently.
The news report states that Qwen-Audio updated its lineup of speech models while reducing the usage prices of related models. The new models cover audio capabilities including understanding, generation, interaction, and creation.
Why it matters to readers
The increase in model types and price reductions may lower the cost for users to try and use Qwen-Audio voice capabilities.
Beginners may focus on use cases such as speech recognition, speech synthesis, and real-time interaction, but the original text does not provide specific integration methods or actual costs.
Risks and unknowns
The original text does not provide specific prices before and after the reduction.
The original text does not provide evaluation results, technical parameters, or real-world performance for the respective models.
The release and price-reduction information in the news is based solely on the supplied report, with no additional verification materials provided.
Related Developments
Loading event timeline…
Related concepts
TTS
This term is not in the glossary yet. Browse related concepts in the glossary.