Pseudo-Labeling for Domain-Agnostic Bangla Automatic Speech Recognition

Rabindra Nath Nandi; Mehadi Menon; Tareq Muntasir; Sagor Sarker; Quazi Sarwar Muhtaseem; Md. Tariqul Islam; Shammur Chowdhury; Firoj Alam

doi:10.18653/v1/2023.banglalp-1.16

Pseudo-Labeling for Domain-Agnostic Bangla Automatic Speech Recognition

Rabindra Nath Nandi, Mehadi Menon, Tareq Muntasir, Sagor Sarker, Quazi Sarwar Muhtaseem, Md. Tariqul Islam, Shammur Chowdhury, Firoj Alam

Abstract

One of the major challenges for developing automatic speech recognition (ASR) for low-resource languages is the limited access to labeled data with domain-specific variations. In this study, we propose a pseudo-labeling approach to develop a large-scale domain-agnostic ASR dataset. With the proposed methodology, we developed a 20k+ hours labeled Bangla speech dataset covering diverse topics, speaking styles, dialects, noisy environments, and conversational scenarios. We then exploited the developed corpus to design a conformer-based ASR system. We benchmarked the trained ASR with publicly available datasets and compared it with other available models. To investigate the efficacy, we designed and developed a human-annotated domain-agnostic test set composed of news, telephony, and conversational data among others. Our results demonstrate the efficacy of the model trained on psuedo-label data for the designed test-set along with publicly-available Bangla datasets. The experimental resources will be publicly available.https://github.com/hishab-nlp/Pseudo-Labeling-for-Domain-Agnostic-Bangla-ASR

Anthology ID:: 2023.banglalp-1.16
Volume:: Proceedings of the First Workshop on Bangla Language Processing (BLP-2023)
Month:: December
Year:: 2023
Address:: Singapore
Editors:: Firoj Alam, Sudipta Kar, Shammur Absar Chowdhury, Farig Sadeque, Ruhul Amin
Venue:: BanglaLP
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 152–162
Language:
URL:: https://aclanthology.org/2023.banglalp-1.16
DOI:: 10.18653/v1/2023.banglalp-1.16
Bibkey:
Cite (ACL):: Rabindra Nath Nandi, Mehadi Menon, Tareq Muntasir, Sagor Sarker, Quazi Sarwar Muhtaseem, Md. Tariqul Islam, Shammur Chowdhury, and Firoj Alam. 2023. Pseudo-Labeling for Domain-Agnostic Bangla Automatic Speech Recognition. In Proceedings of the First Workshop on Bangla Language Processing (BLP-2023), pages 152–162, Singapore. Association for Computational Linguistics.
Cite (Informal):: Pseudo-Labeling for Domain-Agnostic Bangla Automatic Speech Recognition (Nandi et al., BanglaLP 2023)
Copy Citation:
PDF:: https://aclanthology.org/2023.banglalp-1.16.pdf

PDF Cite Search