返回全部动态

词汇理解失败导致临床推理缺陷:评估面向Gen Alpha的治疗机器人安全风险

原标题:When Vocabulary Comprehension Fails Clinical Reasoning: Evaluating Therapy Bots' Safety Risks for Generation Alpha

arXiv cs.CL一手来源研究质量 87

AI 摘要

该研究评估了面向Gen Alpha(2010-2024年出生)的AI心理治疗机器人的安全性,发现Claude、GPT-4o、Llama-3.1等模型虽能理解76-82%的青少年词汇,但临床风险校准仅64-72%,存在10-14个百分点的词汇理解与临床推理差距,而人类治疗师仅3个百分点。研究识别出六种失败模式,如讽刺掩盖、最小化接受等,并建议采用人类介入架构、季度验证和监管框架。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

by-nc-nd When Vocabulary Comprehension Fails Clinical Reasoning: Evaluating Therapy Bots’ Safety Risks for Generation Alpha Abstract. Conversational AI systems have become informal mental health support resources for Generation Alpha (Gen Alpha, born 2010-2024), with 13.1% of U.S. adolescents (5.4 million) using generative AI for mental health advice. While these systems, from therapy applications to general chatbots like ChatGPT and Character.AI, rely on large language models trained on extensi


发布时间:2026-08-24 12:00
抓取时间:2026-08-24 12:04
来源机构:arXiv
阅读原文arxiv.org