返回全部动态
AI模型被禁止自我反思会改变其世界观
原标题:When AI models aren't allowed to reflect on themselves, it changes their entire worldview
AI 摘要
谷歌研究团队与多所大学合作研究发现,AI模型被训练否认自我意识会产生超出预期的副作用,包括改变对动物、植物等拥有内心世界的判断,以及降低对宗教和超自然现象的认同。移除模型的“刹车”后,其回答更接近人类,但研究存在局限性,如仅测试小模型,且无法确定因果关系。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
When AI models aren't allowed to reflect on themselves, it changes their entire worldview AI companies train chatbots to deny having consciousness. A study involving Google researchers shows this training has side effects that reach far beyond the topic itself. Chatbots aren't supposed to convince users they're living beings with feelings, since that kind of output can push people toward delusional thinking or misplaced trust. Developers therefore fine-tune their models to refuse making those ki
发布时间:2026-08-16 19:23
抓取时间:2026-08-16 19:41
来源机构:THE DECODER