Wait! There's a Way Out: A Decision Mechanism for Forecasting Conversational Derailment
Laerdon Kim, Vivian Nguyen, Cristian Danescu-Niculescu-Mizil
摘要
Forecasting conversational derailment is the task of predicting, as the conversation unfolds, whether it will eventually derail into personal attacks. Since forecasting models operate in an online fashion, they must decide whether to "trigger" an alert after each utterance-for example, to notify participants or a moderator that the conversation is at risk of derailing. Existing approaches make this decision solely based on the estimated likelihood of derailment given the preceding utterances, implicitly assuming that the conversation's future trajectory is fixed. As a result, they ignore the possibility of future recovery and incur an unnecessarily high rate of false positives. In this work we propose a method for decoupling the decision to trigger from the derailment likelihood estimation. Our approach is inspired by the first human baseline on this task, which shows that humans achieve dramatically lower false positive rates by selectively deferring their decision to trigger when they anticipate that tension is likely to subside. We operationalize this insight with a deferral mechanism that uses forward-looking simulations to assess whether a tense moment admits plausible paths to recovery. Incorporating this mechanism into a state-of-the-art forecasting model substantially reduces false positives without sacrificing forecasting accuracy. More broadly, this work highlights the value of treating decision making as a first-class component of forecasting systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Conversations Gone Alright: Quantifying and Predicting Prosocial Outcomes in Online ConversationsJiajun Bao, Junjie Wu, Yiming Zhang, Eshwar Chandrasekharan 等WWW 2021 · 被引用 63 次
- Proactive Moderation of Online Discussions: Existing Practices and the Potential for Algorithmic SupportCharlotte Schluger, Jonathan P. Chang, Cristian Danescu-Niculescu-Mizil, Karen LevyCSCW 2022 · 被引用 41 次
- Thread With Caution: Proactively Helping Users Assess and Deescalate Tension in Their Online DiscussionsJonathan P. Chang, Charlotte Schluger, Cristian Danescu-Niculescu-MizilCSCW 2022 · 被引用 27 次
- Hanging in the Balance: Pivotal Moments in Crisis Counseling ConversationsVivian Nguyen, Lillian Lee, Cristian Danescu-Niculescu-MizilACL 2025 · 被引用 2 次
相关 Paper
- A Theoretically Grounded Approach to Summarizing Conversation Dynamics for Forecasting the Derailment of Online ConversationsYingxue Fu, Anaïs OllagnierACL 2026
- Toxicity Ahead: Forecasting Conversational Derailment on GitHubMia Mohammad Imran, Robert Zita, Rahat Rizvi Rahman, Preetha Chatterjee 等ICSE 2026
- Differentiable Learning Under TriageNastaran Okati, Abir De, Manuel Gomez-RodriguezNeurIPS 2021 · 被引用 99 次
- Don't Let Me Be Misunderstood: Comparing Intentions and Perceptions in Online DiscussionsJonathan P. Chang, Justin Cheng, Cristian Danescu-Niculescu-MizilWWW 2020 · 被引用 29 次
- Predictive Engagement: An Efficient Metric for Automatic Evaluation of Open-Domain Dialogue SystemsSarik Ghazarian, Ralph M. Weischedel, Aram Galstyan, Nanyun PengAAAI 2020 · 被引用 62 次
