Temporal and Second Language Influence on Intra-Annotator Agreement and Stability in Hate Speech Labelling

Gavin Abercrombie; Dirk Hovy; Vinodkumar Prabhakaran

doi:10.18653/v1/2023.law-1.10

Temporal and Second Language Influence on Intra-Annotator Agreement and Stability in Hate Speech Labelling

Gavin Abercrombie, Dirk Hovy, Vinodkumar Prabhakaran

Abstract

Much work in natural language processing (NLP) relies on human annotation. The majority of this implicitly assumes that annotator’s labels are temporally stable, although the reality is that human judgements are rarely consistent over time. As a subjective annotation task, hate speech labels depend on annotator’s emotional and moral reactions to the language used to convey the message. Studies in Cognitive Science reveal a ‘foreign language effect’, whereby people take differing moral positions and perceive offensive phrases to be weaker in their second languages. Does this affect annotations as well? We conduct an experiment to investigate the impacts of (1) time and (2) different language conditions (English and German) on measurements of intra-annotator agreement in a hate speech labelling task. While we do not observe the expected lower stability in the different language condition, we find that overall agreement is significantly lower than is implicitly assumed in annotation tasks, which has important implications for dataset reproducibility in NLP.

Anthology ID:: 2023.law-1.10
Volume:: Proceedings of the 17th Linguistic Annotation Workshop (LAW-XVII)
Month:: July
Year:: 2023
Address:: Toronto, Canada
Editors:: Jakob Prange, Annemarie Friedrich
Venue:: LAW
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 96–103
Language:
URL:: https://aclanthology.org/2023.law-1.10
DOI:: 10.18653/v1/2023.law-1.10
Bibkey:
Cite (ACL):: Gavin Abercrombie, Dirk Hovy, and Vinodkumar Prabhakaran. 2023. Temporal and Second Language Influence on Intra-Annotator Agreement and Stability in Hate Speech Labelling. In Proceedings of the 17th Linguistic Annotation Workshop (LAW-XVII), pages 96–103, Toronto, Canada. Association for Computational Linguistics.
Cite (Informal):: Temporal and Second Language Influence on Intra-Annotator Agreement and Stability in Hate Speech Labelling (Abercrombie et al., LAW 2023)
Copy Citation:
PDF:: https://aclanthology.org/2023.law-1.10.pdf
Video:: https://aclanthology.org/2023.law-1.10.mp4

PDF Cite Search Video