Learning Kernel-Smoothed Machine Translation with Retrieved Examples

Qingnan Jiang; Mingxuan Wang; Jun Cao; Shanbo Cheng; Shujian Huang; Lei Li

doi:10.18653/v1/2021.emnlp-main.579

Learning Kernel-Smoothed Machine Translation with Retrieved Examples

Qingnan Jiang, Mingxuan Wang, Jun Cao, Shanbo Cheng, Shujian Huang, Lei Li

Abstract

How to effectively adapt neural machine translation (NMT) models according to emerging cases without retraining? Despite the great success of neural machine translation, updating the deployed models online remains a challenge. Existing non-parametric approaches that retrieve similar examples from a database to guide the translation process are promising but are prone to overfit the retrieved examples. However, non-parametric methods are prone to overfit the retrieved examples. In this work, we propose to learn Kernel-Smoothed Translation with Example Retrieval (KSTER), an effective approach to adapt neural machine translation models online. Experiments on domain adaptation and multi-domain machine translation datasets show that even without expensive retraining, KSTER is able to achieve improvement of 1.1 to 1.5 BLEU scores over the best existing online adaptation methods. The code and trained models are released at https://github.com/jiangqn/KSTER.

Anthology ID:: 2021.emnlp-main.579
Volume:: Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing
Month:: November
Year:: 2021
Address:: Online and Punta Cana, Dominican Republic
Editors:: Marie-Francine Moens, Xuanjing Huang, Lucia Specia, Scott Wen-tau Yih
Venue:: EMNLP
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 7280–7290
Language:
URL:: https://aclanthology.org/2021.emnlp-main.579
DOI:: 10.18653/v1/2021.emnlp-main.579
Bibkey:
Cite (ACL):: Qingnan Jiang, Mingxuan Wang, Jun Cao, Shanbo Cheng, Shujian Huang, and Lei Li. 2021. Learning Kernel-Smoothed Machine Translation with Retrieved Examples. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing, pages 7280–7290, Online and Punta Cana, Dominican Republic. Association for Computational Linguistics.
Cite (Informal):: Learning Kernel-Smoothed Machine Translation with Retrieved Examples (Jiang et al., EMNLP 2021)
Copy Citation:
PDF:: https://aclanthology.org/2021.emnlp-main.579.pdf
Video:: https://aclanthology.org/2021.emnlp-main.579.mp4
Code: jiangqn/kster + additional community code
Data: WMT 2014

PDF Cite Search Code Video