Forgotten Knowledge: Examining the Citational Amnesia in NLP

Janvijay Singh, Mukund Rungta, Diyi Yang, Saif Mohammad


Abstract
Citing papers is the primary method through which modern scientific writing discusses and builds on past work. Collectively, citing a diverse set of papers (in time and area of study) is an indicator of how widely the community is reading. Yet, there is little work looking at broad temporal patterns of citation. This work systematically and empirically examines: How far back in time do we tend to go to cite papers? How has that changed over time, and what factors correlate with this citational attention/amnesia? We chose NLP as our domain of interest and analyzed approximately 71.5K papers to show and quantify several key trends in citation. Notably, around 62% of cited papers are from the immediate five years prior to publication, whereas only about 17% are more than ten years old. Furthermore, we show that the median age and age diversity of cited papers were steadily increasing from 1990 to 2014, but since then, the trend has reversed, and current NLP papers have an all-time low temporal citation diversity. Finally, we show that unlike the 1990s, the highly cited papers in the last decade were also papers with the least citation diversity, likely contributing to the intense (and arguably harmful) recency focus. Code, data, and a demo are available on the project homepage.
Anthology ID:
2023.acl-long.341
Original:
2023.acl-long.341v1
Version 2:
2023.acl-long.341v2
Volume:
Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Month:
July
Year:
2023
Address:
Toronto, Canada
Editors:
Anna Rogers, Jordan Boyd-Graber, Naoaki Okazaki
Venue:
ACL
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
6192–6208
Language:
URL:
https://aclanthology.org/2023.acl-long.341
DOI:
10.18653/v1/2023.acl-long.341
Bibkey:
Cite (ACL):
Janvijay Singh, Mukund Rungta, Diyi Yang, and Saif Mohammad. 2023. Forgotten Knowledge: Examining the Citational Amnesia in NLP. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 6192–6208, Toronto, Canada. Association for Computational Linguistics.
Cite (Informal):
Forgotten Knowledge: Examining the Citational Amnesia in NLP (Singh et al., ACL 2023)
Copy Citation:
PDF:
https://aclanthology.org/2023.acl-long.341.pdf
Video:
 https://aclanthology.org/2023.acl-long.341.mp4