The USTC’s Dialect Speech Translation System for IWSLT 2023

Pan Deng; Shihao Chen; Weitai Zhang; Jie Zhang; Li-Rong Dai

doi:10.18653/v1/2023.iwslt-1.5

The USTC’s Dialect Speech Translation System for IWSLT 2023

Pan Deng, Shihao Chen, Weitai Zhang, Jie Zhang, Lirong Dai

Abstract

This paper presents the USTC system for the IWSLT 2023 Dialectal and Low-resource shared task, which involves translation from Tunisian Arabic to English. We aim to investigate the mutual transfer between Tunisian Arabic and Modern Standard Arabic (MSA) to enhance the performance of speech translation (ST) by following standard pre-training and fine-tuning pipelines. We synthesize a substantial amount of pseudo Tunisian-English paired data using a multi-step pre-training approach. Integrating a Tunisian-MSA translation module into the end-to-end ST model enables the transfer from Tunisian to MSA and facilitates linguistic normalization of the dialect. To increase the robustness of the ST system, we optimize the model’s ability to adapt to ASR errors and propose a model ensemble method. Results indicate that applying the dialect transfer method can increase the BLEU score of dialectal ST. It is shown that the optimal system ensembles both cascaded and end-to-end ST models, achieving BLEU improvements of 2.4 and 2.8 in test1 and test2 sets, respectively, compared to the best published system.

Anthology ID:: 2023.iwslt-1.5
Volume:: Proceedings of the 20th International Conference on Spoken Language Translation (IWSLT 2023)
Month:: July
Year:: 2023
Address:: Toronto, Canada (in-person and online)
Editors:: Elizabeth Salesky, Marcello Federico, Marine Carpuat
Venue:: IWSLT
SIG:: SIGSLT
Publisher:: Association for Computational Linguistics
Note:
Pages:: 102–112
Language:
URL:: https://aclanthology.org/2023.iwslt-1.5/
DOI:: 10.18653/v1/2023.iwslt-1.5
Bibkey:
Cite (ACL):: Pan Deng, Shihao Chen, Weitai Zhang, Jie Zhang, and Lirong Dai. 2023. The USTC’s Dialect Speech Translation System for IWSLT 2023. In Proceedings of the 20th International Conference on Spoken Language Translation (IWSLT 2023), pages 102–112, Toronto, Canada (in-person and online). Association for Computational Linguistics.
Cite (Informal):: The USTC’s Dialect Speech Translation System for IWSLT 2023 (Deng et al., IWSLT 2023)
Copy Citation:
PDF:: https://aclanthology.org/2023.iwslt-1.5.pdf

PDF Cite Search Fix data