Users Hate Blondes: Detecting Sexism in User Comments on Online Romanian News

Andreea Moldovan, Karla Csürös, Ana-maria Bucur, Loredana Bercuci


Abstract
Romania ranks almost last in Europe when it comes to gender equality in political representation, with about 10% fewer women in politics than the E.U. average. We proceed from the assumption that this underrepresentation is also influenced by the sexism and verbal abuse female politicians face in the public sphere, especially in online media. We collect a novel dataset with sexist comments in Romanian language from newspaper articles about Romanian female politicians and propose baseline models using classical machine learning models and fine-tuned pretrained transformer models for the classification of sexist language in the online medium.
Anthology ID:
2022.woah-1.21
Volume:
Proceedings of the Sixth Workshop on Online Abuse and Harms (WOAH)
Month:
July
Year:
2022
Address:
Seattle, Washington (Hybrid)
Editors:
Kanika Narang, Aida Mostafazadeh Davani, Lambert Mathias, Bertie Vidgen, Zeerak Talat
Venue:
WOAH
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
230–230
Language:
URL:
https://aclanthology.org/2022.woah-1.21
DOI:
10.18653/v1/2022.woah-1.21
Bibkey:
Cite (ACL):
Andreea Moldovan, Karla Csürös, Ana-maria Bucur, and Loredana Bercuci. 2022. Users Hate Blondes: Detecting Sexism in User Comments on Online Romanian News. In Proceedings of the Sixth Workshop on Online Abuse and Harms (WOAH), pages 230–230, Seattle, Washington (Hybrid). Association for Computational Linguistics.
Cite (Informal):
Users Hate Blondes: Detecting Sexism in User Comments on Online Romanian News (Moldovan et al., WOAH 2022)
Copy Citation:
PDF:
https://aclanthology.org/2022.woah-1.21.pdf
Video:
 https://aclanthology.org/2022.woah-1.21.mp4