An LLM Feature-based Framework for Dialogue Constructiveness Assessment
Lexin Zhou, Youmna Farag, Andreas Vlachos
摘要
Research on dialogue constructiveness assessment focuses on (i) analysing conversational factors that influence individuals to take specific actions, win debates, change their perspectives or broaden their open-mindedness and (ii) predicting constructiveness outcomes following dialogues for such use cases. These objectives can be achieved by training either interpretable feature-based models (which often involve costly human annotations) or neural models such as pre-trained language models (which have empirically shown higher task accuracy but lack interpretability). In this paper we propose an LLM feature-based framework for dialogue constructiveness assessment that combines the strengths of feature-based and neural approaches, while mitigating their downsides. The framework first defines a set of dataset-independent and interpretable linguistic features, which can be extracted by both prompting an LLM and simple heuristics. Such features are then used to train LLM featurebased models. We apply this framework to three datasets of dialogue constructiveness and find that our LLM feature-based models outperform or performs at least as well as standard feature-based models and neural models. We also find that the LLM feature-based model learns more robust prediction rules instead of relying on superficial shortcuts, which often trouble neural models. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Evaluation and Facilitation of Online Discussions in the LLM Era: A SurveyKaterina Korre, Dimitris Tsirmpas, Nikos Gkoumas, Emma Cabalé 等EMNLP 2025
- Computational Analysis of Conversation Dynamics through Participant ResponsivityMargaret A. Hughes, Brandon Roy, Elinor Poole-Dayan, Deb Roy 等EMNLP 2025
它引用的顶会 Paper4
- Debating with More Persuasive LLMs Leads to More Truthful AnswersAkbir Khan, John Hughes, Dan Valentine, Laura Ruis 等ICML 2024 · 被引用 244 次
- Effects of Persuasive Dialogues: Testing Bot Identities and Inquiry StrategiesWeiyan Shi, Xuewei Wang, Yoojung Oh, Jingwen Zhang 等CHI 2020 · 被引用 93 次
- A Comprehensive Analysis of the Effectiveness of Large Language Models as Automatic Dialogue EvaluatorsChen Zhang, Luis Fernando D'Haro, Yiming Chen, Malu Zhang 等AAAI 2024 · 被引用 57 次
- Towards Argument Mining for Social Good: A SurveyEva Maria Vecchi, Neele Falk, Iman Jundi, Gabriella LapesaACL 2021
相关 Paper
- A Dual-Perspective NLG Meta-Evaluation Framework with Automatic Benchmark and Better InterpretabilityXinyu Hu, Mingqi Gao, Li Lin, Zhenghan Yu 等ACL 2025
- Language Model as an Annotator: Exploring DialoGPT for Dialogue SummarizationXiachong Feng, Xiaocheng Feng, Libo Qin, Bing Qin 等ACL 2021
- Disentangling Language and Culture for Evaluating Multilingual Large Language ModelsJiahao Ying, Wei Tang, Yiran Zhao, Yixin Cao 等ACL 2025 · 被引用 7 次
- Examining Human-AI Collaboration for Co-Writing Constructive Comments OnlineFarhana Shahid, Maximilian Dittgen, Mor Naaman, Aditya VashisthaCSCW 2025 · 被引用 2 次
- X2-DFD: A framework for explainable and extendable Deepfake DetectionYize Chen, Zhiyuan Yan, Guangliang Cheng, Kangran Zhao 等NeurIPS 2025 · 被引用 43 次
