Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models

Chen, Yuyan; Fu, Qiang; Yuan, Yichen; Wen, Zhihao; Fan, Ge; Liu, Dayiheng; Zhang, Dongmei; Li, Zhixu; Xiao, Yanghua

Computer Science > Computation and Language

arXiv:2407.04121v1 (cs)

[Submitted on 4 Jul 2024]

Title:Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models

Authors:Yuyan Chen, Qiang Fu, Yichen Yuan, Zhihao Wen, Ge Fan, Dayiheng Liu, Dongmei Zhang, Zhixu Li, Yanghua Xiao

View PDF HTML (experimental)

Abstract:Large Language Models (LLMs) have gained widespread adoption in various natural language processing tasks, including question answering and dialogue systems. However, a major drawback of LLMs is the issue of hallucination, where they generate unfaithful or inconsistent content that deviates from the input source, leading to severe consequences. In this paper, we propose a robust discriminator named RelD to effectively detect hallucination in LLMs' generated answers. RelD is trained on the constructed RelQA, a bilingual question-answering dialogue dataset along with answers generated by LLMs and a comprehensive set of metrics. Our experimental results demonstrate that the proposed RelD successfully detects hallucination in the answers generated by diverse LLMs. Moreover, it performs well in distinguishing hallucination in LLMs' generated answers from both in-distribution and out-of-distribution datasets. Additionally, we also conduct a thorough analysis of the types of hallucinations that occur and present valuable insights. This research significantly contributes to the detection of reliable answers generated by LLMs and holds noteworthy implications for mitigating hallucination in the future work.

Comments:	Accepted to CIKM 2023 (Long Paper)
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2407.04121 [cs.CL]
	(or arXiv:2407.04121v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2407.04121

Submission history

From: Yuyan Chen [view email]
[v1] Thu, 4 Jul 2024 18:47:42 UTC (977 KB)

Computer Science > Computation and Language

Title:Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators