Bethelhem Mamo

Author directory

2026

Language technologies used in everyday settings such as machine translation systems risk perpetuating societal bias. As previous work shows, biases in these systems not only underscore representational harm but also materialize into economic disparities in resources required to correct errors for the disadvantaged social group. Prior work in creating benchmarks for gender bias in machine translation systems 1) focus primarily on high-resourced language pairs or a low-resourced language paired with a high resource language, 2) use template based benchmarks that usually focus on occupational biases and stereotypes, and 3) translate high-resource benchmarks which may lack cultural significance to low-resourced languages. In this paper, we introduce Yeswa-Stories, a three-way parallel dataset comprising 1,300 aligned sentences in Amharic, Afaan Oromo, and Tigrinya. The dataset focuses on narratives about women and is designed to support research on gender representation in translation. We constructed the dataset in two ways: first, we collected English sentences from Wikipedia articles about notable African women and translated them into the three target languages using human translators. To improve cultural representativeness, we further augment the dataset with locally sourced content reflecting the cultural context where the languages are spoken. Our dataset contributes a new resource for studying gender-inclusive translation in low-resourced settings.