Ask Again, Then Fail: Large Language Models’ Vacillations in Judgment

Qiming Xie, Zengzhi Wang, Yi Feng, Rui Xia


Abstract
We observe that current large language models often waver in their judgments when faced with follow-up questions, even if the original judgment was correct. This wavering presents a significant challenge for generating reliable responses and building user trust. To comprehensively assess this issue, we introduce a Follow-up Questioning Mechanism along with two metrics to quantify this inconsistency, confirming its widespread presence in current large language models. Furthermore, to mitigate this issue, we explore various prompting strategies for closed-source models, and develop a training-based framework Unwavering-FQ that teaches large language models to maintain their originally correct judgments through synthesized high-quality preference data. Our experimental results confirm the effectiveness of our framework and its ability to enhance the general capabilities of large language models.
Anthology ID:
2024.luhme-long.577
Volume:
Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Month:
August
Year:
2024
Address:
Bangkok, Thailand
Editors:
Lun-Wei Ku, Andre Martins, Vivek Srikumar
Venue:
ACL
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
10709–10745
Language:
URL:
https://aclanthology.org/2024.luhme-long.577/
DOI:
10.18653/v1/2024.acl-long.577
Bibkey:
Cite (ACL):
Qiming Xie, Zengzhi Wang, Yi Feng, and Rui Xia. 2024. Ask Again, Then Fail: Large Language Models’ Vacillations in Judgment. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 10709–10745, Bangkok, Thailand. Association for Computational Linguistics.
Cite (Informal):
Ask Again, Then Fail: Large Language Models’ Vacillations in Judgment (Xie et al., ACL 2024)
Copy Citation:
PDF:
https://aclanthology.org/2024.acl-long.577.pdf