To Build Our Future, We Must Know Our Past: Contextualizing Paradigm Shifts in Natural Language Processing
Sireesh Gururaja, Amanda Bertsch, Clara Na, David Gray Widder, Emma Strubell
Abstract
NLP is in a period of disruptive change that is impacting our methodologies, funding sources, and public perception. In this work, we seek to understand how to shape our future by better understanding our past. We study factors that shape NLP as a field, including culture, incentives, and infrastructure by conducting long-form interviews with 26 NLP researchers of varying seniority, research area, institution, and social identity. Our interviewees identify cyclical patterns in the field, as well as new shifts without historical parallel, including changes in benchmark culture and software infrastructure. We complement this discussion with quantitative analysis of citation, authorship, and language use in the ACL Anthology over time. We conclude by discussing shared visions, concerns, and hopes for the future of NLP. We hope that this study of our field’s past and present can prompt informed discussion of our community’s implicit norms and more deliberate action to consciously shape the future.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2c026b86-f5ce-4890-9e5c-de715fc43aaaCited by top-tier papers7
- Power and Play: Investigating "License to Critique" in Teams' AI Ethics DiscussionsDavid Gray Widder, Laura Dabbish, James D. Herbsleb, Nikolas MartelaroCSCW 2024 · 13 citations
- We are Who We Cite: Bridges of Influence Between Natural Language Processing and Other Academic FieldsJan Philip Wahle, Terry Ruas, Mohamed Abdalla, Bela Gipp et al.EMNLP 2023 · 6 citations
- From Insights to Actions: The Impact of Interpretability and Analysis Research on NLPMarius Mosbach, Vagrant Gautam, Tomás Vergara Browne, Dietrich Klakow et al.EMNLP 2024 · 2 citations
- LazyReview: A Dataset for Uncovering Lazy Thinking in NLP Peer ReviewsSukannya Purkayastha, Zhuang Li, Anne Lauscher, Lizhen Qu et al.ACL 2025 · 1 citation
- Research Borderlands: Analysing Writing Across Research CulturesShaily Bhatt, Tal August, Maria AntoniakACL 2025
Builds on8
- "Everyone wants to do the model work, not the data work": Data Cascades in High-Stakes AINithya Sambasivan, Shivani Kapania, Hannah Highfill, Diana Akrong et al.CHI 2021 · 725 citations
- S2ORC: The Semantic Scholar Open Research CorpusKyle Lo, Lucy Lu Wang, Mark Neumann, Rodney Kinney et al.ACL 2020 · 424 citations
- The Elephant in the Room: Analyzing the Presence of Big Tech in Natural Language Processing ResearchMohamed Abdalla, Jan Philip Wahle, Terry Lima Ruas, Aurélie Névéol et al.ACL 2023 · 16 citations
- Geographic Citation Gaps in NLP ResearchMukund Rungta, Janvijay Singh, Saif M. Mohammad, Diyi YangEMNLP 2022 · 10 citations
- Forgotten Knowledge: Examining the Citational Amnesia in NLPJanvijay Singh, Mukund Rungta, Diyi Yang, Saif M. MohammadACL 2023 · 8 citations
Related papers
- What Do NLP Researchers Believe? Results of the NLP Community MetasurveyJulian Michael, Ari Holtzman, Alicia Parrish, Aaron Mueller et al.ACL 2023 · 16 citations
- Examining Citations of Natural Language Processing LiteratureSaif M. MohammadACL 2020
- The ACL OCL Corpus: Advancing Open Science in Computational LinguisticsShaurya Rohatgi, Yanxia Qin, Benjamin Aw, Niranjana Unnithan et al.EMNLP 2023 · 9 citations
- Good Intentions Beyond ACL: Who Does NLP for Social Good, and Where?Grace LeFevre, Qingcheng Zeng, Adam Leif, Jason Jewell et al.EMNLP 2025
- 'I'm Categorizing LLM as a Productivity Tool': Examining Ethics of LLM Use in HCI Research PracticesShivani Kapania, Ruiyi Wang, Toby Jia-Jun Li, Tianshi Li et al.CSCW 2025 · 31 citations
