xAI 低价发布 Grok 4.7,基准测试仍落后 Claude 与 GPT-6
原标题:xAI launches Grok 4.7 at bargain prices, but benchmarks reveal a wide gap to Claude and GPT-6
AI 摘要
xAI 发布其迄今最强模型 Grok 4.7,主打编程与知识工作,基于更大基座模型、更长强化学习训练,并强化自我输出验证。定价为每百万输入 token 2 美元、每百万输出 token 6 美元,接近中国模型价位。但在 Artificial Analysis 智能指数上仅得 46 分,落后于同得 53 分的 Claude Fable 5.1 和 GPT-6;在 Terminal-Bench 4.0 智能体编程测试中仅 26%,远低于 GPT-6 Astra 的 60% 和 Claude Fable 5.1 的 55%,甚至不及 DeepSeek V4.1 Flash 的 27%。模型已通过 Grok API、Cursor 和 Grok Build 提供。
正文节选
xAI launches Grok 4.7 at bargain prices, but benchmarks reveal a wide gap to Claude and GPT-6 Elon Musk's xAI has introduced Grok 4.7, its most capable model yet for coding and knowledge work. It's built on a larger base model, trained with longer reinforcement learning, and designed to better verify its own output, according to the company. Pricing sits at $2 per million input tokens and $6 per million output tokens. Those rates are closer to Chinese models than Western frontier models, probabl