返回全部动态

AI 在这些智力测试中频频失误,你能比它做得更好吗?

原标题:AI models flub these intelligence tests. Can you fare any better?

MIT Technology Review AI研究质量 70

AI 摘要

MIT Technology Review 报道称,AI 模型在空间推理、记忆适应性和抽象视觉推理等智力测试中仍存在明显短板。哥伦比亚大学和谷歌等机构的研究显示,模型在 Connections 谜题上进步显著,但在心理旋转、骑士与无赖变体及 ARC-AGI 等任务上仍常出错。文章通过多个示例谜题,展示了人类与 AI 认知方式的差异,并邀请读者测试自己能否在这些谜题上胜过 AI。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

Puzzles and games have been central to AI development since the very beginning. Just as we humans like to test our smarts with crosswords or logic puzzles, developers can test how far models have advanced with a gaming gauntlet. The term “machine learning” was popularized in a 1959 article by the IBM computer scientist Arthur Samuel about an algorithm that learned to play checkers. Chess and the Chinese board game Go are famous AI test beds too. Judged purely on its puzzling skills, AI is improv


发布时间:2026-08-26 17:00
抓取时间:2026-09-07 03:44
来源机构:MIT Technology Review
阅读原文technologyreview.com