返回全部动态

Ollama v0.33.1 发布:支持 Qwen3.8 Flash Next 与结构化输出

原标题:v0.33.1

Ollama Releases一手来源产品发布质量 68

AI 摘要

Ollama 发布了 v0.33.1 版本,主要更新包括对 MLX 后端 Qwen3.8 Flash Next 的支持,以及 llama.cpp 的更新。此外,mlxrunner 新增了结构化输出支持,并优化了从慢速存储加载模型时的 Metal GPU 超时问题。该版本还修复了 cmake 外部兼容补丁的幂等性问题。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

## What's Changed * MLX: Qwen3.8 Flash Next support * cmake: make external compat patches idempotent * MLX and llama.cpp update * mlxrunner: add structured output support * mlxrunner: avoid Metal GPU timeouts when loading models from slow storage ## New Contributors * @pd95 made their first contribution in https://github.com/ollama/ollama/pull/17948 **Full Changelog**: https://github.com/ollama/ollama/compare/v0.33.0...v0.33.1


发布时间:2026-08-27 02:09
抓取时间:2026-08-28 18:36
来源机构:Ollama
阅读原文github.com