OpenAI新推理技术引发AI安全专家担忧
原标题:OpenAI’s new reasoning technique alarms AI safety experts
AI 摘要
据The Information报道,OpenAI的新模型Astra将采用一种名为“循环深度”(又称“不透明循环”)的推理技术,该技术使模型在循环中处理查询,可能使思维链更难监控,引发AI安全专家的担忧。Redwood Research CEO Buck Shlegeris和AI安全倡导者Zvi Mowshowitz等专家公开表示忧虑,认为该技术可能破坏思维链的可监控性,甚至需要立法干预。OpenAI首席科学家Jakub Pachocki回应称公司致力于保持思维链的可读性,并强调监控是核心研究目标。目前Astra对该技术的使用有限,但Anthropic和Google DeepMind也在讨论类似技术。
正文节选
OpenAI’s new Astra model will use a reasoning technique called “recurrent depth” that allows it to operate outside of the sequential thinking that characterizes most reasoning models, the The Information reported on Tuesday. This technique, also called “opaque recurrence,” will likely make the model’s chain of thought more difficult to monitor — and that has AI safety experts rattled. While Astra’s use of the technique is reportedly limited, its emergence has still raised significant concerns am