HERALD: An Annotation Efficient Method to Detect User Disengagement in Social Conversations
Weixin Liang, Kaihui Liang, Zhou Yu
Abstract
Open-domain dialog systems have a usercentric goal: to provide humans with an engaging conversation experience. User engagement is one of the most important metrics for evaluating open-domain dialog systems, and could also be used as real-time feedback to benefit dialog policy learning. Existing work on detecting user disengagement typically requires hand-labeling many dialog samples. We propose HERALD, an efficient annotation framework that reframes the training data annotation process as a denoising problem. Specifically, instead of manually labeling training samples, we first use a set of labeling heuristics to label training samples automatically. We then denoise the weakly labeled data using the Shapley algorithm. Finally, we use the denoised data to train a user engagement detector. Our experiments show that HERALD improves annotation efficiency significantly and achieves 86% user disengagement detection accuracy in two dialog corpora. Our implementation is available at https:// github.com/Weixin-Liang/HERALD/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- MetaShift: A Dataset of Datasets for Evaluating Contextual Distribution Shifts and Training ConflictsWeixin Liang, James ZouICLR 2022 · 103 citations
- DU-Shapley: A Shapley Value Proxy for Efficient Dataset ValuationFelipe Garrido-Lucero, Benjamin Heymann, Maxime Vono, Patrick Loiseau et al.NeurIPS 2024 · 19 citations
- A Privacy-Friendly Approach to Data ValuationJiachen T. Wang, Yuqing Zhu, Yu-Xiang Wang, Ruoxi Jia et al.NeurIPS 2023 · 12 citations
Builds on6
- Estimating Training Data Influence by Tracing Gradient DescentGarima Pruthi, Frederick Liu, Satyen Kale, Mukund SundararajanNeurIPS 2020 · 784 citations
- Predictive Engagement: An Efficient Metric for Automatic Evaluation of Open-Domain Dialogue SystemsSarik Ghazarian, Ralph M. Weischedel, Aram Galstyan, Nanyun PengAAAI 2020 · 62 citations
- MOSS: End-to-End Dialog System Framework with Modular SupervisionWeixin Liang, Youzhi Tian, Chengcai Chen, Zhou YuAAAI 2020 · 55 citations
- ALICE: Active Learning with Contrastive Natural Language ExplanationsWeixin Liang, James Zou, Zhou YuEMNLP 2020 · 36 citations
- Beyond User Self-Reported Likert Scale Ratings: A Comparison Model for Automatic Dialog EvaluationWeixin Liang, James Zou, Zhou YuACL 2020 · 25 citations
Related papers
- Towards Boosting the Open-Domain Chatbot with Human FeedbackHua Lu, Siqi Bao, Huang He, Fan Wang et al.ACL 2023 · 8 citations
- Data Manipulation: Towards Effective Instance Learning for Neural Dialogue Generation via Learning to Augment and ReweightHengyi Cai, Hongshen Chen, Yonghao Song, Cheng Zhang et al.ACL 2020 · 57 citations
- ADPL: Adversarial Prompt-based Domain Adaptation for Dialogue Summarization with Knowledge DisentanglementLulu Zhao, Fujia Zheng, Weihao Zeng, Keqing He et al.SIGIR 2022 · 6 citations
- Dialogue Response Ranking Training with Large-Scale Human Feedback DataXiang Gao, Yizhe Zhang, Michel Galley, Chris Brockett et al.EMNLP 2020 · 67 citations
- Hierarchical Reinforcement Learning for Open-Domain DialogAbdelrhman Saleh, Natasha Jaques, Asma Ghandeharioun, Judy Hanwen Shen et al.AAAI 2020 · 60 citations
