CodeAnno: Extending WebAnno with Hierarchical Document Level Annotation and Automation

Florian Schneider, Seid Muhie Yimam, Fynn Petersen-frey, Gerret Von Nordheim, Katharina Kleinen-von K”onigsl”ow, Chris Biemann


Abstract
WebAnno is one of the most popular annotation tools that supports generic annotation types and distributive annotation with multiple user roles. However, WebAnno focuses on annotating span-level mentions and relations among them, making document-level annotation complicated. When it comes to the annotation and analysis of social science materials, it usually involves the creation of codes to categorize a given document. The codes, which are known as codebooks, are typically hierarchical, which enables to code the document either with a general category or more fine-grained subcategories. CodeAnno is forked from WebAnno and designed to solve the coding problems faced by many social science researchers with the following main functionalities. 1) Creation of hierarchical codebooks, with functionality to move and sort categories in the hierarchy 2) an interactive UI for codebook annotation 3) import and export of annotations in CSV format, hence being compatible with existing annotations conducted using spreadsheet applications 4) integration of an external automation component to facilitate coding using machine learning 5) project templating that allows duplicating a project structure without copying the actual documents. We present different use-cases to demonstrate the capability of CodeAnno. A shot demonstration video of the system is available here: https://www.youtube.com/watch?v=RmCdTghBe-s
Anthology ID:
2023.eacl-demo.2
Volume:
Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics: System Demonstrations
Month:
May
Year:
2023
Address:
Dubrovnik, Croatia
Editors:
Danilo Croce, Luca Soldaini
Venue:
EACL
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
11–17
Language:
URL:
https://aclanthology.org/2023.eacl-demo.2
DOI:
10.18653/v1/2023.eacl-demo.2
Bibkey:
Cite (ACL):
Florian Schneider, Seid Muhie Yimam, Fynn Petersen-frey, Gerret Von Nordheim, Katharina Kleinen-von K”onigsl”ow, and Chris Biemann. 2023. CodeAnno: Extending WebAnno with Hierarchical Document Level Annotation and Automation. In Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics: System Demonstrations, pages 11–17, Dubrovnik, Croatia. Association for Computational Linguistics.
Cite (Informal):
CodeAnno: Extending WebAnno with Hierarchical Document Level Annotation and Automation (Schneider et al., EACL 2023)
Copy Citation:
PDF:
https://aclanthology.org/2023.eacl-demo.2.pdf
Video:
 https://aclanthology.org/2023.eacl-demo.2.mp4