字节发布全双工语音大模型Seeduplex,已在豆包App上线
原标题:Introducing Seed Full-Duplex Speech LLM: Attentive Listening, Robust Interference Suppression, Enabling More Natural Interaction
AI 摘要
字节跳动旗下Seed研究团队正式推出Seeduplex,一款原生全双工语音大模型,基于全新的“边听边说”框架,支持同时聆听与说话,显著提升交互自然度与流畅度。相比半双工模型,其误响应率和误打断率降低一半,过早响应率降低40%,并已在豆包App全量上线,实现行业首次大规模部署。该模型在复杂声学环境中的干扰抑制和自适应端点检测方面取得突破,多维度评估显示其在对话流畅度和节奏上优于传统方案。
正文节选
Today, we officially introduce Seeduplex, a native full-duplex speech LLM. Compared with the previous-generation Doubao end-to-end speech model based on a half-duplex paradigm, Seeduplex is built on an entirely new "listen while speaking" framework, significantly enhancing the naturalness and fluency of the interaction experience, delivering a notable leap in the naturalness and fluency of interactive experiences. If an end-to-end architecture, by unifying the "listening" and "speaking" modules,