ALW: Adaptive Layer-Wise contrastive decoding enhancing reasoning ability in Large Language Models

Yuechi Zhou; Chuyue Zhou; Jianxin Zhang; Juntao Li; Min Zhang

doi:10.18653/v1/2025.findings-acl.447

ALW: Adaptive Layer-Wise contrastive decoding enhancing reasoning ability in Large Language Models

Yuechi Zhou, Chuyue Zhou, Jianxin Zhang, Juntao Li, Min Zhang

Abstract

Large language models (LLMs) have achieved remarkable performance across various reasoning tasks. However, many LLMs still encounter challenges in reasoning, especially for LLMs with fewer parameters or insufficient pre-training data. Through our experiments, we identify that noise accumulation across layers often leads to unstable token predictions during reasoning. We find that contrasting the probability distributions across layers effectively mitigates this interference. Building on this insight, we propose Adaptive Layer-Wise contrastive decoding (ALW), a novel framework that enhances reasoning ability by dynamically disentangling noise in shallow layers from critical signals in deep layers. Extensive experiments on several reasoning benchmarks demonstrate that ALW consistently improves answer accuracy across multiple LLMs while maintaining inference efficiency. For example, we achieve a 48% improvement on the Gsm8k using the LLaMA-7B model and an absolute accuracy increase of 5.2 points on the BBH evaluation benchmark with the LLaMA-65B model.

Anthology ID:: 2025.findings-acl.447
Volume:: Findings of the Association for Computational Linguistics: ACL 2025
Month:: July
Year:: 2025
Address:: Vienna, Austria
Editors:: Wanxiang Che, Joyce Nabende, Ekaterina Shutova, Mohammad Taher Pilehvar
Venue:: Findings
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 8506–8524
Language:
URL:: https://aclanthology.org/2025.findings-acl.447/
DOI:: 10.18653/v1/2025.findings-acl.447
Bibkey:
Cite (ACL):: Yuechi Zhou, Chuyue Zhou, Jianxin Zhang, Juntao Li, and Min Zhang. 2025. ALW: Adaptive Layer-Wise contrastive decoding enhancing reasoning ability in Large Language Models. In Findings of the Association for Computational Linguistics: ACL 2025, pages 8506–8524, Vienna, Austria. Association for Computational Linguistics.
Cite (Informal):: ALW: Adaptive Layer-Wise contrastive decoding enhancing reasoning ability in Large Language Models (Zhou et al., Findings 2025)
Copy Citation:
PDF:: https://aclanthology.org/2025.findings-acl.447.pdf

PDF Cite Search Fix data