AgentionAI 发布 Qwen3.8-27B 高保真 GGUF 量化
原标题:agentionai/Qwen3.8-27B-AP-GGUF
AI 摘要
AgentionAI 在 Hugging Face 发布 Qwen3.8-27B 的 Agention Precision GGUF 量化包,基于 Qwen/Qwen3.8-27B 量化,采用非均匀分配与低损失纠错编码,兼容标准 llama.cpp 无需分支或额外参数,并包含视觉投影器和 MTP 草稿头。作者称在相同文件大小、速度和显存下,各档位量化相比 Unsloth 等公开 GGUF 更接近 BF16 原始模型的下一 token 分布,并公开了 KL 散度评测协议与复现命令。
正文节选
--- base_model: - Qwen/Qwen3.8-27B base_model_relation: quantized license: apache-2.0 library_name: gguf pipeline_tag: image-text-to-text tags: - gguf - qwen3.8-27b - imatrix - agentionai --- # Qwen3.8-27B · Agention Precision GGUF [](https://github.com/sponsors/LaurentZuijdwijk) **Same size. Same speed. Closer to the model Qwen trained than any other quant** ![Lower KL at the same G