返回全部动态
Mistral AI 发布 Voxtral 语音模型,开源且成本减半
原标题:ResearchVoxtralJuly 15, 2025By Mistral AI
AI 摘要
Mistral AI 发布了 Voxtral 语音理解模型,提供 24B 和 3B 两种规模,均采用 Apache 2.0 许可证,并可通过 API 使用。该模型支持长上下文、多语言、问答、摘要和函数调用,在转录和音频理解基准上优于 Whisper 等模型,且成本低于同类 API。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
Voice: the original UI. Voice was humanity’s first interface—long before writing or typing, it let us share ideas, coordinate work, and build relationships. As digital systems become more capable, voice is returning as our most natural form of human-computer interaction. Yet today’s systems remain limited—unreliable, proprietary, and too brittle for real-world use. Closing this gap demands tools with exceptional transcription, deep understanding, multilingual fluency, and open, flexible deployme
发布时间:—
抓取时间:2026-08-02 00:32
来源机构:Mistral AI