Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent

AI Summary
Alibaba's AI team Qwen has released Qwen-Audio-3. 1, a lineup of five models for speech recognition (ASR), text-to-speech (TTS), and real-time interaction.
From the source
Alibaba's AI team Qwen has released Qwen-Audio-3.1, a lineup of five models for speech recognition (ASR), text-to-speech (TTS), and real-time interaction. The ASR model improves multilingual and dialect recognition and automatically cleans up filler words and repetitions. ASR-Next adds multi-speaker identification with timestamps and detects emotions, ambient sounds, and machine noise. TTS handles multilingual synthesis […] The article Alibaba launches Qwen Audio 3.1 with new models and slashes
The full text couldn't be loaded here (the source may require a subscription).
View original at The Decoder