返回全部动态
Ollama v0.32.3 发布:修复下载与工具调用,扩展 GPU 支持
原标题:v0.32.3
AI 摘要
Ollama 发布 v0.32.3 版本,修复了模型下载停滞、GLM 工具调用丢失等问题,并改进了 Claude Code 通道和 Anthropic 思考流等集成。该版本扩展了 GPU 支持,包括 Windows ARM64 上的 CUDA、通过 CUDA 12 支持 B200,并降低了 Linux 上 CUDA/ROCm iGPU 的内存使用。此外,为 Laguna 2.1 模型增加了聊天、思考和工具调用支持,并更新了 MLX 和 llama.cpp 引擎。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
## What's Changed - Fixed model downloads that stall before sending data. - Improved integrations: restored Claude Code Channels, fixed Anthropic thinking streams, and made Hermes Desktop respect `--force-build`. - Expanded GPU support with CUDA on Windows ARM64, B200 support through CUDA 12, and lower memory use on Linux CUDA/ROCm iGPUs. - Added chat, thinking, and tool calling support for Laguna 2.1 models, including a Metal inference fix. - Fixed GLM tool calls being silently dropped a
发布时间:2026-07-23 08:44
抓取时间:2026-08-02 00:27
来源机构:Ollama