返回全部动态

bartowski 发布 DeepSeek-V4-Flash-0731 的 MXFP4 量化版 GGUF

原标题:bartowski/DeepSeek-V4-Flash-0731-GGUF

Hugging Face New and Trending Models一手来源模型发布质量 67

AI 摘要

bartowski 使用 llama.cpp b10173 对 deepseek-ai 的 DeepSeek-V4-Flash-0731 模型进行了量化,发布了 GGUF 格式的 MXFP4 版本,文件大小为 156.38GB。该模型拥有 284B 参数,目前仅提供 MXFP4 格式,其他量化大小可能后续推出。用户可通过 llama.cpp、LM Studio 等工具运行此量化模型。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

--- quantized_by: bartowski pipeline_tag: text-generation license: mit base_model_relation: quantized base_model: deepseek-ai/DeepSeek-V4-Flash-0731 --- ## Llamacpp Quantizations of DeepSeek-V4-Flash-0731 by deepseek-ai Using <a href="https://github.com/ggml-org/llama.cpp/">llama.cpp</a> release <a href="https://github.com/ggml-org/llama.cpp/releases/tag/b10173">b10173</a> for quantization. Original model: https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731 This model is in MXFP4 and a


发布时间:2026-08-05 01:53
抓取时间:2026-08-05 01:55
来源机构:Hugging Face
阅读原文huggingface.co