Too Consistent to Detect: A Study of Self-Consistent Errors in LLMs

Hexiang Tan; Fei Sun; Sha Liu; Du Su; Qi Cao; Xin Chen (陈鑫); Jingang Wang; Xunliang Cai; Yuanzhuo Wang; Huawei Shen (沈华伟); Xueqi Cheng (程学旗)

doi:10.18653/v1/2025.emnlp-main.238

Too Consistent to Detect: A Study of Self-Consistent Errors in LLMs

Hexiang Tan, Fei Sun, Sha Liu, Du Su, Qi Cao, Xin Chen, Jingang Wang, Xunliang Cai, Yuanzhuo Wang, Huawei Shen, Xueqi Cheng

Abstract

As large language models (LLMs) often generate plausible but incorrect content, error detection has become increasingly critical to ensure truthfulness.However, existing detection methods often overlook a critical problem we term as **self-consistent error**, where LLMs repeatedly generate the same incorrect response across multiple stochastic samples.This work formally defines self-consistent errors and evaluates mainstream detection methods on them.Our investigation reveals two key findings: (1) Unlike inconsistent errors, whose frequency diminishes significantly as the LLM scale increases, the frequency of self-consistent errors remains stable or even increases.(2) All four types of detection methods significantly struggle to detect self-consistent errors.These findings reveal critical limitations in current detection methods and underscore the need for improvement.Motivated by the observation that self-consistent errors often differ across LLMs, we propose a simple but effective cross‐model probe method that fuses hidden state evidence from an external verifier LLM.Our method significantly enhances performance on self-consistent errors across three LLM families.

Anthology ID:: 2025.emnlp-main.238
Volume:: Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing
Month:: November
Year:: 2025
Address:: Suzhou, China
Editors:: Christos Christodoulopoulos, Tanmoy Chakraborty, Carolyn Rose, Violet Peng
Venue:: EMNLP
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 4755–4765
Language:
URL:: https://aclanthology.org/2025.emnlp-main.238/
DOI:: 10.18653/v1/2025.emnlp-main.238
Bibkey:
Cite (ACL):: Hexiang Tan, Fei Sun, Sha Liu, Du Su, Qi Cao, Xin Chen, Jingang Wang, Xunliang Cai, Yuanzhuo Wang, Huawei Shen, and Xueqi Cheng. 2025. Too Consistent to Detect: A Study of Self-Consistent Errors in LLMs. In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, pages 4755–4765, Suzhou, China. Association for Computational Linguistics.
Cite (Informal):: Too Consistent to Detect: A Study of Self-Consistent Errors in LLMs (Tan et al., EMNLP 2025)
Copy Citation:
PDF:: https://aclanthology.org/2025.emnlp-main.238.pdf
Checklist:: 2025.emnlp-main.238.checklist.pdf

PDF Cite Search Checklist Fix data