返回全部动态

AtomicChat 发布 DeepSeek-V4-Flash GGUF 量化版,质量对比领先

原标题:AtomicChat/DeepSeek-V4-Flash-0731-GGUF

Hugging Face New and Trending Models一手来源开源质量 85

AI 摘要

AtomicChat 发布了 DeepSeek-V4-Flash-0731 的 GGUF 量化版本,该模型为 284B 参数的 MoE 模型,已进行量化感知训练,专家权重以 MXFP4 存储。文章提供了详细的量化质量对比表,显示在相同大小下其量化版本优于 unsloth 的版本,并警告 GGUF 文件中的聊天模板已过时,需手动替换以避免推理能力下降。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

--- license: mit base_model: deepseek-ai/DeepSeek-V4-Flash-0731 base_model_relation: quantized quantized_by: AtomicChat pipeline_tag: text-generation library_name: gguf tags: - gguf - llama.cpp - deepseek - deepseek-v4 - deepseek_v4 - moe - mixture-of-experts - imatrix - quantized - mxfp4 - qat - 1-bit - 2-bit - 3-bit - conversational - atomic-chat language: - en - zh metrics: - perplexity - kl_divergence --- # DeepSeek-V4-Flash-0731-GGUF > [!WARNING] > **The chat template inside these GGUF fi


发布时间:2026-08-05 03:54
抓取时间:2026-08-05 02:30
来源机构:Hugging Face
阅读原文huggingface.co