SNaC: Coherence Error Detection for Narrative Summarization

Tanya Goyal; Junyi Jessy Li; Greg Durrett

doi:10.18653/v1/2022.emnlp-main.29

SNaC: Coherence Error Detection for Narrative Summarization

Tanya Goyal, Junyi Jessy Li, Greg Durrett

Abstract

Progress in summarizing long texts is inhibited by the lack of appropriate evaluation frameworks. A long summary that appropriately covers the facets of that text must also present a coherent narrative, but current automatic and human evaluation methods fail to identify gaps in coherence. In this work, we introduce SNaC, a narrative coherence evaluation framework for fine-grained annotations of long summaries. We develop a taxonomy of coherence errors in generated narrative summaries and collect span-level annotations for 6.6k sentences across 150 book and movie summaries. Our work provides the first characterization of coherence errors generated by state-of-the-art summarization models and a protocol for eliciting coherence judgments from crowdworkers. Furthermore, we show that the collected annotations allow us to benchmark past work in coherence modeling and train a strong classifier for automatically localizing coherence errors in generated summaries. Finally, our SNaC framework can support future work in long document summarization and coherence evaluation, including improved summarization modeling and post-hoc summary correction.

Anthology ID:: 2022.emnlp-main.29
Volume:: Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing
Month:: December
Year:: 2022
Address:: Abu Dhabi, United Arab Emirates
Editors:: Yoav Goldberg, Zornitsa Kozareva, Yue Zhang
Venue:: EMNLP
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 444–463
Language:
URL:: https://aclanthology.org/2022.emnlp-main.29/
DOI:: 10.18653/v1/2022.emnlp-main.29
Bibkey:
Cite (ACL):: Tanya Goyal, Junyi Jessy Li, and Greg Durrett. 2022. SNaC: Coherence Error Detection for Narrative Summarization. In Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing, pages 444–463, Abu Dhabi, United Arab Emirates. Association for Computational Linguistics.
Cite (Informal):: SNaC: Coherence Error Detection for Narrative Summarization (Goyal et al., EMNLP 2022)
Copy Citation:
PDF:: https://aclanthology.org/2022.emnlp-main.29.pdf
Video:: https://aclanthology.org/2022.emnlp-main.29.mp4

PDF Cite Search Video Fix data