返回全部动态

Ollama v0.32.1 发布:改进 Gemma 4 工具调用与 MLX 缓存

原标题:v0.32.1

Ollama Releases一手来源产品发布质量 72

AI 摘要

Ollama 发布 v0.32.1 版本,改进了 Gemma 4 的工具调用和多轮推理能力,修复了 MLX 模型缓存泄漏问题,并优化了缓存快照性能。此外,该版本还支持在 MLX 文本模型加载时遵守 OLLAMA_LOAD_TIMEOUT 设置,并增强了代理的 Web 搜索和获取功能,提示用户进行身份验证。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

## What's Changed - Improved Gemma 4 tool calling and multi-turn reasoning, including more reliable tool-response continuations - Fixed a recurrent MLX model cache leak that could increase memory use across requests, and improved cache snapshot performance - MLX text model loading now respects `OLLAMA_LOAD_TIMEOUT` - Agent web search and fetch now tell users to run `ollama signin` when authentication is required - The interactive agent now receives the current working directory for better p


发布时间:2026-07-16 11:27
抓取时间:2026-08-02 00:27
来源机构:Ollama
阅读原文github.com