返回全部动态

阿里发布 Qwen Audio 3.1 并最高降价 95%

原标题:Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent

THE DECODER模型发布质量 62

AI 摘要

阿里巴巴通义千问团队发布 Qwen-Audio-3.1,包含五款覆盖语音识别(ASR)、文本转语音(TTS)和实时交互的模型。ASR 模型提升多语言与方言识别能力,可自动清理口头语和重复内容;ASR-Next 新增带时间戳的多说话人识别,并能检测情绪、环境音和机器噪声。TTS 支持多语言合成与跨语言音色迁移,用户可通过文本提示控制情绪、语速和风格;TTS-Next 结合语言模型与扩散方法一次性生成语音、音效和背景音;实时模型支持边说边听与即时打断。同时阿里大幅降价,TTS 约降 70%,Realtime 约降 85%,ASR 最高降 95%。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent Alibaba's AI team Qwen has released Qwen-Audio-3.1, a lineup of five models for speech recognition (ASR), text-to-speech (TTS), and real-time interaction. The ASR model improves multilingual and dialect recognition and automatically cleans up filler words and repetitions. ASR-Next adds multi-speaker identification with timestamps and detects emotions, ambient sounds, and machine noise. TTS handles mul


发布时间:2026-09-23 20:31
抓取时间:2026-09-23 21:01
来源机构:THE DECODER
阅读原文the-decoder.com