Skip to main content
QUICK REVIEW

[论文解读] AI Hallucinations: A Misnomer Worth Clarifying

Negar Maleki, Balaji Padmanabhan|arXiv (Cornell University)|Jan 9, 2024
Artificial Intelligence in Healthcare and Education被引用 5
一句话总结

该论文对跨越14个数据库的AI幻觉定义进行了系统性综述,揭示不存在普遍定义,并提出标准术语和分类法。

ABSTRACT

As large language models continue to advance in Artificial Intelligence (AI), text generation systems have been shown to suffer from a problematic phenomenon termed often as "hallucination." However, with AI's increasing presence across various domains including medicine, concerns have arisen regarding the use of the term itself. In this study, we conducted a systematic review to identify papers defining "AI hallucination" across fourteen databases. We present and analyze definitions obtained across all databases, categorize them based on their applications, and extract key points within each category. Our results highlight a lack of consistency in how the term is used, but also help identify several alternative terms in the literature. We discuss implications of these and call for a more unified effort to bring consistency to an important contemporary AI issue that can affect multiple domains significantly.

研究动机与目标

  • 识别在不同领域和数据库中,AI幻觉这一术语的定义方式。
  • 评估定义的一致性并识别共同特征与替代表述。
  • 汇编并分类定义,以为AI生成内容错误提供统一术语。
  • 为跨领域使用提供走向正式定义和健壮分类法的指南。

提出的方法

  • 在2013年至2023年之间,对14个数据库进行广泛文献检索,寻找在AI/LLMs中对AI幻觉的定义论文。
  • 对每篇检索到的论文进行人工审核,提取定义及背景信息。
  • 汇总333条定义并在附录中总结。
  • 按应用领域对定义进行分类(如健康、法律、翻译、摘要等)。
  • 识别替代表述并讨论对术语标准化的影响。

实验结果

研究问题

  • RQ1在不同领域和数据库中,存在哪些对AI幻觉的定义?
  • RQ2这些定义的一致性如何,它们具有哪些共同特征或存在的分歧?
  • RQ3使用了哪些替代表述,如何制定统一的分类法以提高清晰度?

主要发现

  • 在文献中没有一个精确的、被普遍接受的AI幻觉定义。
  • 定义因应用而异,可能存在冲突或依赖于情境。
  • 表II和表III概述了跨领域描述AI幻觉时使用的替代术语及要点。
  • 作者汇编了2013–2023年的333条定义,并在附录中提供完整定义集。
  • 近来文献中有推动替换或重新命名该术语的趋势,以避免与心理健康的含义和污名相关联。
  • 论文倡导统一术语和健全、正式的定义,以改善跨领域的沟通与研究。

更好的研究,从现在开始

从阅读论文到最终审阅,大幅缩短您的研究时间。

无需绑定信用卡

本解读由 AI 生成,并经人工编辑审核。