返回全部动态
Ollama v0.33.1 发布:支持 Qwen3.8 Flash Next 与结构化输出
原标题:v0.33.1
AI 摘要
Ollama 发布了 v0.33.1 版本,主要更新包括对 MLX 后端 Qwen3.8 Flash Next 的支持,以及 llama.cpp 的更新。此外,mlxrunner 新增了结构化输出支持,并优化了从慢速存储加载模型时的 Metal GPU 超时问题。该版本还修复了 cmake 外部兼容补丁的幂等性问题。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
## What's Changed * MLX: Qwen3.8 Flash Next support * cmake: make external compat patches idempotent * MLX and llama.cpp update * mlxrunner: add structured output support * mlxrunner: avoid Metal GPU timeouts when loading models from slow storage ## New Contributors * @pd95 made their first contribution in https://github.com/ollama/ollama/pull/17948 **Full Changelog**: https://github.com/ollama/ollama/compare/v0.33.0...v0.33.1
发布时间:2026-08-27 02:09
抓取时间:2026-08-28 18:36
来源机构:Ollama