返回全部动态

Claude Fable 5.1 生成精美动画鹈鹕

原标题:Claude Fable 5.1 made me a really nice animated pelican

Simon Willison's Weblog产品发布质量 75

AI 摘要

Anthropic 发布了 Claude Fable 5.1 模型,声称在编码、知识工作和长期问题解决方面树立了新标准,并在新的 Terminal-Bench-Science 基准上取得 52.6% 的分数。作者 Simon Willison 测试了该模型生成鹈鹕骑自行车 SVG 的能力,发现低和中等推理级别下模型几乎不进行推理,而高和最高级别则产生更详细和高质量的结果,最高级别花费 13 分钟和 3.30 美元。作者还将最高级别的输出动画化,尽管车轮旋转方向错误,但整体动画效果良好。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

Claude Fable 5.1 made me a really nice animated pelican 1st September 2026 Today is Claude Fable (and Mythos) 5.1 day. Anthropic say that Fable 5.1 “sets a new standard for coding, knowledge work, and long-running problem-solving tasks”. Their announcement spends a notable amount of time on scientific research, boasting of a 52.6% score on the brand new Terminal-Bench-Science 0.1 benchmark (first announced on August 27th), up from 24.7% for Fable 5, 29.0% for Opus 5 and 22.4% for GPT-5.6 Sol. Ot


发布时间:2026-09-02 07:57
抓取时间:2026-09-02 08:12
来源机构:Simon Willison
阅读原文simonwillison.net