返回全部动态

Qwen3.8-27B 无审查模型发布六档 NVFP4 GGUF 量化版

原标题:esatapedico/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NVFP4-GGUF

Hugging Face New and Trending Models一手来源模型发布质量 58

AI 摘要

Hugging Face 用户 esatapedico 发布了 DavidAU 的 Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored 模型的六个 GGUF 量化版本,采用 NVFP4(W4A16,FP8 缩放,组大小 16)格式,面向 Blackwell sm_120 架构。六个文件共享字节完全一致的 448 张量 NVFP4 主干,仅在 lm_head、token_embd 和 MTP 头的精度上形成从 VERY-LOW 到 VERY-HIGH 的阶梯,体积从 14.86 GB 到 19.69 GB。每个文件内置 MTP 投机解码头,无需单独 drafter,并附带 2026-09-12 与 09-15 两次多模态/工具调用模板修复。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

--- license: apache-2.0 base_model: DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored pipeline_tag: text-generation library_name: gguf description: "Six GGUFs of DavidAU's Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored (TWIN-TURBO Fable Cold Fusion Heretic Uncensored, 709 arc-c class): NVFP4 family with a per-tier quality ladder for lm_head / token_embd / MTP head. All six tiers share a byte-identical 448-tensor native-NVFP4 backbone. MTP speculative decoding baked int


发布时间:2026-09-16 04:14
抓取时间:2026-09-16 04:14
来源机构:Hugging Face
阅读原文huggingface.co