A Representation Level Analysis of NMT Model Robustness to Grammatical Errors

Abderrahmane Issam; Yusuf Can Semerci; Jan Scholtes; Gerasimos Spanakis

doi:10.18653/v1/2025.findings-acl.451

A Representation Level Analysis of NMT Model Robustness to Grammatical Errors

Abderrahmane Issam, Yusuf Can Semerci, Jan Scholtes, Gerasimos Spanakis

Abstract

Understanding robustness is essential for building reliable NLP systems. Unfortunately, in the context of machine translation, previous work mainly focused on documenting robustness failures or improving robustness. In contrast, we study robustness from a model representation perspective by looking at internal model representations of ungrammatical inputs and how they evolve through model layers. For this purpose, we perform Grammatical Error Detection (GED) probing and representational similarity analysis. Our findings indicate that the encoder first detects the grammatical error, then corrects it by moving its representation toward the correct form. To understand what contributes to this process, we turn to the attention mechanism where we identify what we term *Robustness Heads*. We find that *Robustness Heads* attend to interpretable linguistic units when responding to grammatical errors, and that when we fine-tune models for robustness, they tend to rely more on *Robustness Heads* for updating the ungrammatical word representation.

Anthology ID:: 2025.findings-acl.451
Volume:: Findings of the Association for Computational Linguistics: ACL 2025
Month:: July
Year:: 2025
Address:: Vienna, Austria
Editors:: Wanxiang Che, Joyce Nabende, Ekaterina Shutova, Mohammad Taher Pilehvar
Venue:: Findings
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 8579–8601
Language:
URL:: https://aclanthology.org/2025.findings-acl.451/
DOI:: 10.18653/v1/2025.findings-acl.451
Bibkey:
Cite (ACL):: Abderrahmane Issam, Yusuf Can Semerci, Jan Scholtes, and Gerasimos Spanakis. 2025. A Representation Level Analysis of NMT Model Robustness to Grammatical Errors. In Findings of the Association for Computational Linguistics: ACL 2025, pages 8579–8601, Vienna, Austria. Association for Computational Linguistics.
Cite (Informal):: A Representation Level Analysis of NMT Model Robustness to Grammatical Errors (Issam et al., Findings 2025)
Copy Citation:
PDF:: https://aclanthology.org/2025.findings-acl.451.pdf

PDF Cite Search Fix data