Argument Summarization and its Evaluation in the Era of Large Language Models
Moritz Altemeyer, Steffen Eger, Johannes Daxenberger, Yanran Chen, Tim Altendorf, Philipp Cimiano, Benjamin Schiller
2025Year
3Top-tier citations
Abstract
Moritz Altemeyer, Steffen Eger, Johannes Daxenberger, Yanran Chen, Tim Altendorf, Philipp Cimiano, Benjamin Schiller. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. 2025.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 337552cd-ac19-4a4e-8299-121678e3e295Cited by top-tier papers3
- From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judgeDawei Li, Bohan Jiang, Liangjie Huang, Alimohammad Beigi et al.EMNLP 2025 · 37 citations
- Arg-LLaDA: Argument Summarization via Large Language Diffusion Models and Sufficiency-Aware RefinementHao Li, Yizheng Sun, Viktor Schlegel, Kailai Yang et al.ACL 2026 · 1 citation
- ArgCMV: An Argument Summarization Benchmark for the LLM-eraOmkar Gurjar, Agam Goyal, Eshwar ChandrasekharanEMNLP 2025
Builds on10
- BARTScore: Evaluating Generated Text as Text GenerationWeizhe Yuan, Graham Neubig, Pengfei LiuNeurIPS 2021 · 1,143 citations
- G-Eval: NLG Evaluation using Gpt-4 with Better Human AlignmentYang Liu, Dan Iter, Yichong Xu, Shuohang Wang et al.EMNLP 2023 · 549 citations
- PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning OptimizationYidong Wang, Zhuohao Yu, Wenjin Yao, Zhengran Zeng et al.ICLR 2024 · 368 citations
- BLEURT: Learning Robust Metrics for Text GenerationThibault Sellam, Dipanjan Das, Ankur P. ParikhACL 2020 · 40 citations
- INSTRUCTSCORE: Towards Explainable Text Generation Evaluation with Automatic FeedbackWenda Xu, Danqing Wang, Liangming Pan, Zhenqiao Song et al.EMNLP 2023 · 36 citations
Related papers
- PychoAgent: Psychology-driven LLM Agents for Explainable Panic Prediction on Social Media during Sudden Disaster EventsMengzhu Liu, Zhengqiu Zhu, Chuan Ai, Chen Gao et al.EMNLP 2025 · 1 citation
- TeleMelody: Lyric-to-Melody Generation with a Template-Based Two-Stage MethodZeqian Ju, Peiling Lu, Xu Tan, Rui Wang et al.EMNLP 2022 · 19 citations
- Counter Turing Test (CT2): AI-Generated Text Detection is Not as Easy as You May Think - Introducing AI Detectability Index (ADI)Megha Chakraborty, S. M. Towhidul Islam Tonmoy, S. M. Mehedi Zaman, Shreya Gautam et al.EMNLP 2023 · 14 citations
- EvoWiki: Evaluating LLMs on Evolving KnowledgeWei Tang, Yixin Cao, Yang Deng, Jiahao Ying et al.ACL 2025
- Compare to The Knowledge: Graph Neural Fake News Detection with External KnowledgeLinmei Hu, Tianchi Yang, Luhao Zhang, Wanjun Zhong et al.ACL 2021
