ArgFuse: A Weakly-Supervised Framework for Document-Level Event Argument Aggregation

Debanjana Kar, Sudeshna Sarkar, Pawan Goyal


Abstract
Most of the existing information extraction frameworks (Wadden et al., 2019; Veysehet al., 2020) focus on sentence-level tasks and are hardly able to capture the consolidated information from a given document. In our endeavour to generate precise document-level information frames from lengthy textual records, we introduce the task of Information Aggregation or Argument Aggregation. More specifically, our aim is to filter irrelevant and redundant argument mentions that were extracted at a sentence level and render a document level information frame. Majority of the existing works have been observed to resolve related tasks of document-level event argument extraction (Yang et al., 2018; Zheng et al., 2019) and salient entity identification (Jain et al., 2020) using supervised techniques. To remove dependency from large amounts of labelled data, we explore the task of information aggregation using weakly supervised techniques. In particular, we present an extractive algorithm with multiple sieves which adopts active learning strategies to work efficiently in low-resource settings. For this task, we have annotated our own test dataset comprising of 131 document information frames and have released the code and dataset to further research prospects in this new domain. To the best of our knowledge, we are the first to establish baseline results for this task in English. Our data and code are publicly available at https://github.com/DebanjanaKar/ArgFuse.
Anthology ID:
2021.case-1.5
Volume:
Proceedings of the 4th Workshop on Challenges and Applications of Automated Extraction of Socio-political Events from Text (CASE 2021)
Month:
August
Year:
2021
Address:
Online
Editor:
Ali Hürriyetoğlu
Venue:
CASE
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
20–30
Language:
URL:
https://aclanthology.org/2021.case-1.5
DOI:
10.18653/v1/2021.case-1.5
Bibkey:
Cite (ACL):
Debanjana Kar, Sudeshna Sarkar, and Pawan Goyal. 2021. ArgFuse: A Weakly-Supervised Framework for Document-Level Event Argument Aggregation. In Proceedings of the 4th Workshop on Challenges and Applications of Automated Extraction of Socio-political Events from Text (CASE 2021), pages 20–30, Online. Association for Computational Linguistics.
Cite (Informal):
ArgFuse: A Weakly-Supervised Framework for Document-Level Event Argument Aggregation (Kar et al., CASE 2021)
Copy Citation:
PDF:
https://aclanthology.org/2021.case-1.5.pdf
Video:
 https://aclanthology.org/2021.case-1.5.mp4
Code
 DebanjanaKar/ArgFuse
Data
SciREX