A Multi-persona Framework for Argument Quality Assessment
Bojun Jin, Jianzhu Bao, Yufang Hou, Yang Sun, Yice Zhang, Huajie Wang, Bin Liang, Ruifeng Xu
摘要
Argument quality assessment faces inherent challenges due to its subjective nature, where different evaluators may assign varying quality scores for an argument based on personal perspectives. Although existing datasets collect opinions from multiple annotators to model subjectivity, most existing computational methods fail to consider multi-perspective evaluation. To address this issue, we propose MPAQ, a multi-persona framework for argument quality assessment that simulates diverse evaluator perspectives through large language models. It first dynamically generates targeted personas tailored to an input argument, then simulates each persona's reasoning process to evaluate the argument quality from multiple perspectives. To effectively generate fine-grained quality scores, we develop a coarse-to-fine scoring strategy that first generates a coarse-grained integer score and then refines it into a finegrained decimal score. Experiments on IBM-Rank-30k and IBM-ArgQ-5.3kArgs datasets demonstrate that MPAQ consistently outperforms strong baselines while providing comprehensive multi-perspective rationales.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- A Large-Scale Dataset for Argument Quality Ranking: Construction and AnalysisShai Gretz, Roni Friedman, Edo Cohen-Karlik, Assaf Toledo 等AAAI 2020 · 被引用 148 次
- Efficient Pairwise Annotation of Argument QualityLukas Gienapp, Benno Stein, Matthias Hagen, Martin PotthastACL 2020 · 被引用 17 次
- ArgAnalysis35K : A large-scale dataset for Argument Quality AnalysisOmkar Joshi, Priya Pitre, Yashodhara HaribhaktaACL 2023 · 被引用 4 次
- Argue with Me Tersely: Towards Sentence-Level Counter-Argument GenerationJiayu Lin, Rong Ye, Meng Han, Qi Zhang 等EMNLP 2023 · 被引用 2 次
- Architectural Sweet Spots for Modeling Human Label Variation by the Example of Argument Quality: It's Best to Relate Perspectives!Philipp Heinisch, Matthias Orlikowski, Julia Romberg, Philipp CimianoEMNLP 2023 · 被引用 1 次
相关 Paper
- Multi-Agent-as-Judge: Aligning LLM-Agent-Based Automated Evaluation with Multi-Dimensional Human EvaluationJiaju Chen, Yuxuan Lu, Xiaojie Wang, Huimin Zeng 等ACL 2026 · 被引用 30 次
- Contextual Interaction for Argument Post Quality AssessmentYiran Wang, Xuanang Chen, Ben He, Le SunEMNLP 2023 · 被引用 4 次
- Let's discuss! Quality Dimensions and Annotated Datasets for Computational Argument Quality AssessmentRositsa V. Ivanova, Thomas Huber, Christina NiklausEMNLP 2024 · 被引用 2 次
- Hitting your MARQ: Multimodal ARgument Quality Assessment in Long Debate VideoMd. Kamrul Hasan, James Spann, Masum Hasan, Md. Saiful Islam 等EMNLP 2021
- Towards Multi-dimensional Evaluation of LLM Summarization across Domains and LanguagesHyangsuk Min, Yuho Lee, Minjeong Ban, Jiaqi Deng 等ACL 2025 · 被引用 8 次
