返回全部动态
Qwen3.8-Flash-Next:多模态 MoE 模型,Qwen4 架构预览
原标题:Qwen3.8-Flash-Next
AI 摘要
Qwen 发布了新的开源权重模型 Qwen3.8-Flash-Next,这是一个多模态 MoE 模型,也是 Qwen4 架构的早期预览。模型总参数为 125B,但仅有 6B 激活,性能显著提升。作者 Simon Willison 在 DGX Spark 上使用 Unsloth 量化版本进行了测试,并分享了初步体验。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
26th August 2026 - Link Blog Qwen3.8-Flash-Next (via) Another open weights model from Qwen. This one is "a multimodal MoE model that also serves as an early preview of the architecture used in Qwen4". It's pretty big: 125B tokens, but only 6B active which means it gets a significant performance boost. I've been trying it out on a DGX Spark using these Unsloth quantized models. I'm still exploring the model - so far I've tried the 72.5GB UD-IQ1_S one (producing these pelicans) and the 78.9GB UD-Q
发布时间:2026-08-27 07:52
抓取时间:2026-08-28 10:34
来源机构:Simon Willison