Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent
5 Articles
5 Articles
The price of some models in this series has been reduced by up to 95%. The post Alibaba unveils Qwen-Audio-3.1 models for voice understanding and real-time conversation appeared first on Digiato.
Alibaba expanded its Qwen platform with five models specialized in voice recognition, synthesis and real-time conversation, while announcing reductions of up to 95% in their prices. The update also incorporates speaker identification, emotion detection and style control through text instructions.
Alibaba's AI team Qwen has introduced Qwen Audio 3.1, a comprehensive audio model series that includes five models for speech recognition (ASR), speech synthesis (TTS) and real-time interaction. In voice recognition, the ASR model offers improved multilingual recognition and dialect recognition as well as automatic cleaning of fill words and repetitions. The new ASR-Next also recognizes multiple speakers with timestamps and understands emotions …
Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent
Alibaba's AI team Qwen has released Qwen-Audio-3.1, a lineup of five models for speech recognition (ASR), text-to-speech (TTS), and real-time interaction. The ASR model improves multilingual and dialect recognition and automatically cleans up filler words and repetitions. ASR-Next adds multi-speaker identification with timestamps and detects emotions, ambient sounds, and machine noise. TTS handles multilingual synthesis […] The article Alibaba l…
Alibaba Upgrades Qwen Audio Suite: LiveTranslate Latency Cut and Qwen-Audio-3.1 ASR/TTS/Realtime
Alibaba introduced a fresh multimodal speech and audio lineup at the 2026 Apsara Conference, led by Qwen3.8-LiveTranslate and a Qwen-Audio-3.1 product family, according to Alibaba Cloud's English press release dated September 22. Official English branding is Alibaba / Qwen. The angle here is audio product specifications—latency, recognition, synthesis and realtime interaction—not training roadmaps or agent-cloud infrastructure from the same week…
Coverage Details
Bias Distribution
- There is no tracked Bias information for the sources covering this story.
Factuality
To view factuality data please Upgrade to Premium




