返回全部动态

Mistral AI 发布 Voxtral TTS 多语言语音生成模型

原标题:Voxtral TTS

Mistral AI News一手来源模型发布质量 79

AI 摘要

Mistral AI 发布了 Voxtral TTS,一个 4B 参数的多语言文本转语音模型,支持 9 种语言,具有低延迟和情感表达。该模型可通过 API 和 Mistral Studio 使用,定价为每 1k 字符 0.016 美元,并支持零样本跨语言语音适应。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

Thinking Summary Mistral AI has launched Voxtral TTS, a 4B-parameter text-to-speech model delivering state-of-the-art multilingual voice generation with realistic, emotionally expressive speech in 9 languages, low latency, and easy voice adaptation. The model excels in contextual understanding, speaker modeling, and zero-shot cross-lingual voice adaptation, making it ideal for enterprise voice workflows and scalable AI agents. Available via API and Mistral Studio, it offers cost-effective, high-


发布时间:
抓取时间:2026-08-02 00:22
来源机构:Mistral AI
阅读原文mistral.ai