%0 Conference Proceedings %T Sentence Concatenation Approach to Data Augmentation for Neural Machine Translation %A Kondo, Seiichiro %A Hotate, Kengo %A Hirasawa, Tosho %A Kaneko, Masahiro %A Komachi, Mamoru %Y Durmus, Esin %Y Gupta, Vivek %Y Liu, Nelson %Y Peng, Nanyun %Y Su, Yu %S Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Student Research Workshop %D 2021 %8 June %I Association for Computational Linguistics %C Online %F kondo-etal-2021-sentence %X Recently, neural machine translation is widely used for its high translation accuracy, but it is also known to show poor performance at long sentence translation. Besides, this tendency appears prominently for low resource languages. We assume that these problems are caused by long sentences being few in the train data. Therefore, we propose a data augmentation method for handling long sentences. Our method is simple; we only use given parallel corpora as train data and generate long sentences by concatenating two sentences. Based on our experiments, we confirm improvements in long sentence translation by proposed data augmentation despite the simplicity. Moreover, the proposed method improves translation quality more when combined with back-translation. %R 10.18653/v1/2021.naacl-srw.18 %U https://aclanthology.org/2021.naacl-srw.18 %U https://doi.org/10.18653/v1/2021.naacl-srw.18 %P 143-149