Reproducibility in Computational Linguistics: Is Source Code Enough?
Mohammad Arvan, Luís Pina, Natalie Parde
摘要
The availability of source code has been put forward as one of the most critical factors for improving the reproducibility of scientific research. This work studies trends in source code availability at major computational linguistics conferences, namely, ACL, EMNLP, LREC, NAACL, and COLING. We observe positive trends, especially in conferences that actively promote reproducibility. We follow this by conducting a reproducibility study of eight papers published in EMNLP 2021, finding that source code releases leave much to be desired. Moving forward, we suggest all conferences require self-contained artifacts and provide a venue to evaluate such artifacts at the time of publication. Authors can include small-scale experiments and explicit scripts to generate each result to improve the reproducibility of their work.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Beyond the Leaderboard: Rethinking Medical Benchmarks for Large Language ModelsWenxuan Wang, Zizhan Ma, Guo Yu, Yiu-Fai Cheung 等ACL 2026 · 被引用 9 次
- The Unreasonable Effectiveness of Open Science in AI: A Replication StudyOdd Erik Gundersen, Odd Cappelen, Martin Mølnå, Nicklas Grimstad NilsenAAAI 2025 · 被引用 8 次
- We Need to Talk About Reproducibility in NLP Model ComparisonYan Xue, Xuefei Cao, Xingli Yang, Yu Wang 等EMNLP 2023 · 被引用 2 次
- Forest vs Tree: The (N, K) Trade-off in Reproducible ML EvaluationDeepak Pandita, Flip Korn, Chris Welty, Christopher M. HomanAAAI 2026 · 被引用 2 次
- GSAP-ERE: Fine-Grained Scholarly Entity and Relation Extraction Focused on Machine LearningWolfgang Otto, Lu Gan, Sharmila Upadhyaya, Saurav Karmakar 等AAAI 2026
它引用的顶会 Paper8
- Weakly-supervised Text Classification Based on Keyword GraphLu Zhang, Jiandong Ding, Yi Xu, Yingyao Liu 等EMNLP 2021 · 被引用 46 次
- StreamHover: Livestream Transcript Summarization and AnnotationSangwoo Cho, Franck Dernoncourt, Tim Ganter, Trung Bui 等EMNLP 2021 · 被引用 18 次
- Measuring Association Between Labels and Free-Text RationalesSarah Wiegreffe, Ana Marasovic, Noah A. SmithEMNLP 2021 · 被引用 12 次
- Automatically Exposing Problems with Neural Dialog ModelsDian Yu, Kenji SagaeEMNLP 2021 · 被引用 5 次
- A Massively Multilingual Analysis of Cross-linguality in Shared Embedding SpaceAlexander Jones, William Yang Wang, Kyle MahowaldEMNLP 2021 · 被引用 4 次
相关 Paper
- NLP Reproducibility For All: Understanding Experiences of BeginnersShane Storks, Keunwoo Peter Yu, Ziqiao Ma, Joyce ChaiACL 2023
- Code replicability in computer graphicsNicolas Bonneel, David Coeurjolly, Julie Digne, Nicolas MelladoSIGGRAPH 2020 · 被引用 21 次
- "Get in Researchers; We're Measuring Reproducibility": A Reproducibility Study of Machine Learning Papers in Tier 1 Security ConferencesDaniel Olszewski, Allison Lu, Carson Stillman, Kevin Warren 等CCS 2023 · 被引用 19 次
- Reflections on the Reproducibility of Commercial LLM Performance in Empirical Software Engineering StudiesFlorian Angermeir, Maximilian Amougou, Mark Kreitz, Andreas Bauer 等ICSE 2026 · 被引用 1 次
- The State of Open Science in Software Engineering Research: A Case Study of ICSE ArtifactsAl Muttakin, Saikat Mondal, Chanchal K. RoyICSE 2026
