返回全部动态
llama.cpp b10576 发布:重新引入 SYCL Q2_K 内核
原标题:b10576
AI 摘要
llama.cpp 发布 b10576 版本,重新引入了 SYCL 后端针对 Q2_K 量化格式的重排序 MMVQ 和 ESIMD 内核,并添加了门控参数。该版本提供了适用于 macOS、Linux、Windows、Android 和 iOS 的多种预编译二进制文件,支持 CPU、Vulkan、CUDA、ROCm、OpenVINO、SYCL 等后端。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
<details open> sycl : add Q2_K reordered MMVQ and ESIMD kernels (again) (#27490) * Revert "Revert "sycl : add Q2_K reordered MMVQ and ESIMD kernels (#26336)" (#…" This reverts commit 7a0e42fd01fb0acda644e4f04b1f1acbbb9e23ba. * add gate params </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/42300392> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b10576/llama-b1057
发布时间:2026-08-22 15:49
抓取时间:2026-08-22 16:15
来源机构:ggml-org