Information Locality as an Inductive Bias for Neural Language Models

Taiga Someya; Anej Svete; Brian DuSell; Timothy O’Donnell; Mario Giulianelli; Ryan Cotterell

doi:10.18653/v1/2025.acl-long.1357

Information Locality as an Inductive Bias for Neural Language Models

Taiga Someya, Anej Svete, Brian DuSell, Timothy J. O’Donnell, Mario Giulianelli, Ryan Cotterell

Abstract

Inductive biases are inherent in every machine learning system, shaping how models generalize from finite data. In the case of neural language models (LMs), debates persist as to whether these biases align with or diverge from human processing constraints. To address this issue, we propose a quantitative framework that allows for controlled investigations into the nature of these biases. Within our framework, we introduce m-local entropy—an information-theoretic measure derived from average lossy-context surprisal—that captures the local uncertainty of a language by quantifying how effectively the preceding symbols disambiguate the next symbol. In experiments on both perturbed natural language corpora and languages defined by probabilistic finite-state automata (PFSA), we show that languages with higher m-local entropy are more difficult for Transformer and LSTM LMs to learn. These results suggest that neural LMs, much like humans, are highly sensitive to the local statistical structure of a language.

Anthology ID:: 2025.acl-long.1357
Volume:: Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Month:: July
Year:: 2025
Address:: Vienna, Austria
Editors:: Wanxiang Che, Joyce Nabende, Ekaterina Shutova, Mohammad Taher Pilehvar
Venue:: ACL
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 27995–28013
Language:
URL:: https://aclanthology.org/2025.acl-long.1357/
DOI:: 10.18653/v1/2025.acl-long.1357
Bibkey:
Cite (ACL):: Taiga Someya, Anej Svete, Brian DuSell, Timothy J. O’Donnell, Mario Giulianelli, and Ryan Cotterell. 2025. Information Locality as an Inductive Bias for Neural Language Models. In Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 27995–28013, Vienna, Austria. Association for Computational Linguistics.
Cite (Informal):: Information Locality as an Inductive Bias for Neural Language Models (Someya et al., ACL 2025)
Copy Citation:
PDF:: https://aclanthology.org/2025.acl-long.1357.pdf

PDF Cite Search Fix data