返回全部动态
Qwen3.8-27B NVFP4 MTP GGUF 发布:八种精度文件支持视觉与投机解码
原标题:esatapedico/Qwen3.8-27B-NVFP4-MTP-GGUF
AI 摘要
esatapedico 发布了 Qwen3.8-27B-NVFP4-MTP-GGUF,包含八个 GGUF 文件,针对 Blackwell 硬件优化,保留了 NVFP4 密度并内置 MTP 投机解码头。该模型是原生视觉语言模型,支持图像和视频,原生上下文长度 262,144。不同文件在精度和大小上有所区分,以适应不同显存需求。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
--- license: apache-2.0 base_model: unsloth/Qwen3.8-27B-NVFP4 pipeline_tag: text-generation library_name: gguf description: "Family of 8 GGUFs of Qwen3.8-27B (NVFP4): ORIG (source-preserving NVFP4 MLP + BF16 attention), VERY-LOW / LOW-MINUS / LOW / MEDIUM / HIGH / VERY-HIGH (448-tensor byte-identical NVFP4 backbone + per-tier heads), HIGHEST (NVFP4 MLP + Q8_0 attention/lm_head + BF16 embeddings). Native VLM (vision+video), MTP speculative decoding, 262,144 native context. Blackwell sm_120." tags
发布时间:2026-08-19 01:30
抓取时间:2026-08-19 01:30
来源机构:Hugging Face