返回全部动态

llama.cpp b10505 发布:新增 dedup-cache-models 预设选项

原标题:b10505

llama.cpp Releases一手来源产品发布质量 63

AI 摘要

llama.cpp 发布 b10505 版本,为 server 新增 dedup-cache-models 预设选项(PR #27346),用于模型缓存去重。该版本提供了覆盖 macOS、Linux、Windows、Android 及 openEuler 等多平台的预编译二进制文件,支持 CPU、Vulkan、CUDA、ROCm、OpenVINO、SYCL 等多种后端。部分平台构建(如 macOS KleidiAI、Ubuntu ROCm、openEuler)因兼容性问题被禁用。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

<details open> server: add dedup-cache-models preset option (#27346) </details> **Website:** - <https://llama.app> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b10505/llama-b10505-bin-macos-arm64.tar.gz) - macOS Apple Silicon (arm64, KleidiAI enabled) [DISABLED](https://github.com/ggml-org/llama.cpp/pull/23780) - [macOS Intel (x64)](https://github.com/ggml-org/llama.cpp/releases/download/b10505/llama-b10505-bin-macos-x64.tar.gz) - [iO


发布时间:2026-08-20 10:36
抓取时间:2026-08-20 11:27
来源机构:ggml-org
阅读原文github.com