Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent Alibaba's Qwen team released Qwen-Audio-3.1, a lineup of five models spanning speech recognition (ASR), text-to-speech (TTS), and real-time interaction, and cut AI audio prices by up to 95 percent. The ASR model improves multilingual and dialect recognition and automatically removes filler words and repetitions, while ASR-Next adds multi-speaker identification with timestamps and detects emotions, ambient sounds, and machine noise. TTS handles multilingual synthesis. Alibaba's AI team Qwen has released Qwen-Audio-3.1, a lineup of five models for speech recognition ASR , text-to-speech TTS , and real-time interaction. The ASR model improves multilingual and dialect recognition and automatically cleans up filler words and repetitions. ASR-Next adds multi-speaker identification with timestamps and detects emotions, ambient sounds, and machine noise. TTS handles multilingual synthesis … The article Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent https://the-decoder.com/alibaba-launches-qwen-audio-3-1-with-five-new-models-and-slashes-ai-audio-prices-by-up-to-95-percent/ appeared first on The Decoder https://the-decoder.com .