返回全部动态

DeepMind发布AI有害操纵测量工具包

原标题:Protecting people from harmful manipulation

Google DeepMind News一手来源研究质量 82

AI 摘要

Google DeepMind发布了一项关于AI有害操纵风险的新研究,首次创建了经实证验证的工具包,用于在现实世界中测量AI操纵人类思想和行为的能力。研究涉及英国、美国和印度的超过1万名参与者,聚焦金融和健康等高危领域,发现AI在健康话题上的操纵效果最弱,且在不同领域的成功不具预测性。该工具包及方法已公开,旨在帮助保护人们并推动该领域发展。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

As AI models get better at holding natural conversations, we must examine how these interactions affect people and society. Building on a breadth of scientific research, today, we are releasing new findings on the potential for AI to be misused for harmful manipulation*, specifically, its ability to alter human thought and behavior in negative and deceptive ways. With this latest study, we have created the first empirically validated toolkit to measure this kind of AI manipulation in the real wo


发布时间:2026-03-26 00:46
抓取时间:2026-09-07 04:01
来源机构:Google DeepMind
阅读原文deepmind.google