uir-cis at SemEval-2025 Task 3: Detection of Hallucinations in Generated Text

Jia Huang; Shuli Zhao; Yaru Zhao; Tao Chen; Weijia Zhao; Hangui Lin; Yiyang Chen; Binyang Li

uir-cis at SemEval-2025 Task 3: Detection of Hallucinations in Generated Text

Jia Huang, Shuli Zhao, Yaru Zhao, Tao Chen, Weijia Zhao, Hangui Lin, Yiyang Chen, Binyang Li

Abstract

The widespread deployment of large language models (LLMs) across diverse domains has underscored the critical need to ensure the credibility and accuracy of their generated content, particularly in the presence of hallucinations. These hallucinations can severely compromise both the practical performance of models and the security of their applications. In response to this issue, SemEval-2025 Task 3 Mu-SHROOM: Multilingual Shared-task on Hallucinations and Related Observable Overgeneration Mistakes introduces a more granular task for hallucination detection. This task seeks to identify hallucinations in text, accurately locate hallucinated segments, and assess their credibility. In this paper, we present a three-stage method for fine-grained hallucination detection and localization. First, we transform the text into a triplet representation, facilitating more precise hallucination analysis. Next, we leverage a large language model to generate fact-reference texts that correspond to the triplets. Finally, we employ a fact alignment strategy to identify and localize hallucinated segments by evaluating the semantic consistency between the extracted triplets and the generated reference texts. We evaluate our method on the unlabelled test set across all languages in Task 3, demonstrating strong detection performance and validating its effectiveness in multilingual contexts.

Anthology ID:: 2025.semeval-1.134
Volume:: Proceedings of the 19th International Workshop on Semantic Evaluation (SemEval-2025)
Month:: July
Year:: 2025
Address:: Vienna, Austria
Editors:: Sara Rosenthal, Aiala Rosá, Debanjan Ghosh, Marcos Zampieri
Venues:: SemEval | WS
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 1015–1022
Language:
URL:: https://aclanthology.org/2025.semeval-1.134/
DOI:
Bibkey:
Cite (ACL):: Jia Huang, Shuli Zhao, Yaru Zhao, Tao Chen, Weijia Zhao, Hangui Lin, Yiyang Chen, and Binyang Li. 2025. uir-cis at SemEval-2025 Task 3: Detection of Hallucinations in Generated Text. In Proceedings of the 19th International Workshop on Semantic Evaluation (SemEval-2025), pages 1015–1022, Vienna, Austria. Association for Computational Linguistics.
Cite (Informal):: uir-cis at SemEval-2025 Task 3: Detection of Hallucinations in Generated Text (Huang et al., SemEval 2025)
Copy Citation:
PDF:: https://aclanthology.org/2025.semeval-1.134.pdf

PDF Cite Search Fix data