礼貌的人工智能减轻用户对AI幻觉的易感性

Polite AI mitigates user susceptibility to AI hallucinations

Ergonomics · 2024
被引 11 · 同刊同年前 6%
ABS 3

中文导读

通过实验发现,使用礼貌语气的人工智能聊天助手能帮助用户更敏锐地识别AI生成的错误信息(幻觉),并采取更谨慎的判断策略,对高风险场景有实际意义。

Abstract

With their increased capability, AI-based chatbots have become increasingly popular tools to help users answer complex queries. However, these chatbots may hallucinate, or generate incorrect but very plausible-sounding information, more frequently than previously thought. Thus, it is crucial to examine strategies to mitigate human susceptibility to hallucinated output. In a between-subjects experiment, participants completed a difficult quiz with assistance from either a polite or neutral-toned AI chatbot, which occasionally provided hallucinated (incorrect) information. Signal detection analysis revealed that participants interacting with polite-AI showed modestly higher sensitivity in detecting hallucinations and a more conservative response bias compared to those interacting with neutral-toned AI. While the observed effect sizes were modest, even small improvements in users' ability to detect AI hallucinations can have significant consequences, particularly in high-stakes domains or when aggregated across millions of AI interactions.

人工智能自然语言处理心理学人机交互