COGUMELO at SemEval-2025 Task 3: A Synthetic Approach to Detecting Hallucinations in Language Models based on Named Entity Recognition

Aldan Creo; Héctor Cerezo - Costas; Maximiliano Hormazábal Lagos; Pedro Alonso Doval

COGUMELO at SemEval-2025 Task 3: A Synthetic Approach to Detecting Hallucinations in Language Models based on Named Entity Recognition

Aldan Creo, Héctor Cerezo - Costas, Maximiliano Hormazábal Lagos, Pedro Alonso Doval

Abstract

In this paper, we propose an approach to detecting hallucinations based on a Named Entity Recognition (NER) task.We focus on efficiency, aiming to develop a model that can detect hallucinations without relying on external data sources or expensive computations that involve state-of-the-art large language models with upwards of tens of billions of parameters. We utilize the SQuAD question answering dataset to generate a synthetic version that contains both correct and hallucinated responses and train encoder language models of a moderate size (RoBERTa and FLAN-T5) to predict spans of text that are highly likely to contain a hallucination. We test our models on a separate dataset of expert-annotated question-answer pairs and find that our approach achieves a Jaccard similarity of up to 0.358 and 0.227 Spearman correlation, which suggests that our models can serve as moderately accurate hallucination detectors, ideally as part of a detection pipeline involving human supervision. We also observe that larger models seem to develop an emergent ability to leverage their background knowledge to make more informed decisions, while smaller models seem to take shortcuts that can lead to a higher number of false positives.We make our data and code publicly accessible, along with an online visualizer. We also release our trained models under an open license.

Anthology ID:: 2025.semeval-1.281
Volume:: Proceedings of the 19th International Workshop on Semantic Evaluation (SemEval-2025)
Month:: July
Year:: 2025
Address:: Vienna, Austria
Editors:: Sara Rosenthal, Aiala Rosá, Debanjan Ghosh, Marcos Zampieri
Venues:: SemEval | WS
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 2170–2176
Language:
URL:: https://aclanthology.org/2025.semeval-1.281/
DOI:
Bibkey:
Cite (ACL):: Aldan Creo, Héctor Cerezo - Costas, Maximiliano Hormazábal Lagos, and Pedro Alonso Doval. 2025. COGUMELO at SemEval-2025 Task 3: A Synthetic Approach to Detecting Hallucinations in Language Models based on Named Entity Recognition. In Proceedings of the 19th International Workshop on Semantic Evaluation (SemEval-2025), pages 2170–2176, Vienna, Austria. Association for Computational Linguistics.
Cite (Informal):: COGUMELO at SemEval-2025 Task 3: A Synthetic Approach to Detecting Hallucinations in Language Models based on Named Entity Recognition (Creo et al., SemEval 2025)
Copy Citation:
PDF:: https://aclanthology.org/2025.semeval-1.281.pdf

PDF Cite Search Fix data