返回全部动态

llama.cpp b10359 发布:修复 WebGPU CI 错误并支持多平台

原标题:b10359

llama.cpp Releases一手来源产品发布质量 63

AI 摘要

llama.cpp 发布 b10359 版本,主要修复了 ggml-webgpu 的 CI 错误,并添加了 i32 支持到 cpy 操作。该版本提供了适用于 macOS、Linux、Windows、Android 等多个平台的预编译二进制文件,涵盖 CPU、Vulkan、CUDA、ROCm 等多种后端。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

<details open> ggml-webgpu: fix CI errors from #25025 and #25262 (#26566) * test new flash_attn test * rebase and fix to disable subgrou matrices when max_kv_tile == 0 * delete log output * Add i32 support to cpy and enables the all ops test * restore the non target ci tests * comment out of TODO of build-cpu.yml * fix format </details> **Website:** - <https://llama.app> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b10359/llama


发布时间:2026-08-11 18:48
抓取时间:2026-08-12 01:46
来源机构:ggml-org
阅读原文github.com