产品

Alibaba 发布 Qwen Audio 3.1 系列五款语音模型,AI 音频服务降价最高 95%

Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent

精选理由

Qwen 一次发了五款语音模型,能识别方言、区分多个说话人还能标注情绪,价格还砍了 95%,做语音应用的可以看看。

Alibaba 旗下 Qwen 团队发布 Qwen-Audio-3.1,包含五款覆盖语音识别(ASR)、语音合成(TTS)和实时交互的模型。ASR 模型改进了多语言和方言识别,能自动清理语气词和重复内容。新增的 ASR-Next 支持多说话人区分并带时间戳,可检测情绪、环境音和机器噪声。TTS 模型支持多语言语音合成。同时 Alibaba 将 AI 音频服务价格下调最高 95%。

原文 · Decoder

Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent

Alibaba's AI team Qwen has released Qwen-Audio-3.1, a lineup of five models for speech recognition (ASR), text-to-speech (TTS), and real-time interaction. The ASR model improves multilingual and dialect recognition and automatically cleans up filler words and repetitions. ASR-Next adds multi-speaker identification with timestamps and detects emotions, ambient sounds, and machine noise. TTS handles multilingual synthesis […] The article Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent appeared first on The Decoder .