Sense and Sensitivity: Examining the Influence of Semantic Recall on Long Context Code Understanding

Adam Štorek; Mukur Gupta; Samira Hajizadeh; Prashast Srivastava; Suman Jana

Sense and Sensitivity: Examining the Influence of Semantic Recall on Long Context Code Understanding

Adam Štorek, Mukur Gupta, Samira Hajizadeh, Prashast Srivastava, Suman Jana

Abstract

Large language models (LLMs) are increasingly deployed for understanding large codebases, but whether they understand operational semantics of long code context or rely on pattern matching shortcuts remains unclear. We distinguish between lexical recall (retrieving code verbatim) and semantic recall (understanding operational semantics). Evaluating 10 state-of-the-art LLMs, we find that while frontier models achieve near-perfect, position-independent lexical recall, semantic recall degrades severely when code is centrally positioned in long contexts. We introduce semantic recall sensitivity to measure whether tasks require understanding of code’s operational semantics vs. permit pattern matching shortcuts. Through a novel counterfactual measurement method, we show that models rely heavily on pattern matching shortcuts to solve existing code understanding benchmarks. We propose a new task SemTrace, which achieves high semantic recall sensitivity through unpredictable operations; LLMs’ accuracy exhibits severe positional effects, with median accuracy drops of 92.73% versus CRUXEval’s 53.36% as the relevant code snippet approaches the middle of the input code context. Our findings suggest current evaluations substantially underestimate semantic recall failures in long context code understanding.

Anthology ID:: 2026.acl-long.19
Volume:: Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Month:: July
Year:: 2026
Address:: San Diego, California, United States
Editors:: Maria Liakata, Viviane P. Moreira, Jiajun Zhang, David Jurgens
Venue:: ACL
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 480–498
Language:
URL:: https://aclanthology.org/2026.acl-long.19/
DOI:
Bibkey:
Cite (ACL):: Adam Štorek, Mukur Gupta, Samira Hajizadeh, Prashast Srivastava, and Suman Jana. 2026. Sense and Sensitivity: Examining the Influence of Semantic Recall on Long Context Code Understanding. In Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 480–498, San Diego, California, United States. Association for Computational Linguistics.
Cite (Informal):: Sense and Sensitivity: Examining the Influence of Semantic Recall on Long Context Code Understanding (Štorek et al., ACL 2026)
Copy Citation:
PDF:: https://aclanthology.org/2026.acl-long.19.pdf
Checklist:: 2026.acl-long.19.checklist.pdf

PDF Cite Search Checklist Fix data