Examining Human-AI Collaboration for Co-Writing Constructive Comments Online
Farhana Shahid, Maximilian Dittgen, Mor Naaman, Aditya Vashistha
摘要
This paper examines if large language models (LLMs) can help people write constructive comments on divisive social issues due to the difficulty of expressing constructive disagreement online. Through controlled experiments with 600 participants from India and the US, who reviewed and wrote constructive comments on threads related to Islamophobia and homophobia, we observed potential misalignment between how LLMs and humans perceive constructiveness in online comments. While the LLM was more likely to prioritize politeness and balance among contrasting viewpoints when evaluating constructiveness, participants emphasized logic and facts more than the LLM did. Despite these differences, participants rated both LLM-generated and human-AI co-written comments as significantly more constructive than those written independently by humans. Our analysis also revealed that LLM-generated comments integrated significantly more linguistic features of constructiveness compared to human-written comments. When participants used LLMs to refine their comments, the resulting comments were more constructive, more positive, less toxic, and retained the original intent. However, LLMs often distorted people's original views-especially when their stances were on a spectrum instead of being outright polarizing. Based on these findings, we discuss ethical and design considerations in using LLMs to facilitate constructive discourse online.
CCS Concepts: • Human-centered computing → Empirical studies in HCI .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Governance of AI-Generated Content: A Case Study on Social Media PlatformsLan Gao, Abani Ahmed, Oscar Chen, Margaux Reyl 等CHI 2026 · 被引用 3 次
- LLMs Homogenize Values in Constructive Arguments on Value-Laden TopicsFarhana Shahid, Stella Zhang, Aditya VashisthaCHI 2026 · 被引用 1 次
- Evaluation and Facilitation of Online Discussions in the LLM Era: A SurveyKaterina Korre, Dimitris Tsirmpas, Nikos Gkoumas, Emma Cabalé 等EMNLP 2025
- SimulatorArena: Are User Simulators Reliable Proxies for Multi-Turn Evaluation of AI Assistants?Yao Dou, Michel Galley, Baolin Peng, Chris Kedzie 等EMNLP 2025
它引用的顶会 Paper23
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- CoAuthor: Designing a Human-AI Collaborative Writing Dataset for Exploring Language Model CapabilitiesMina Lee, Percy Liang, Qian YangCHI 2022 · 被引用 340 次
- Co-Writing with Opinionated Language Models Affects Users' ViewsMaurice Jakesch, Advait Bhat, Daniel Buschek, Lior Zalmanson 等CHI 2023 · 被引用 249 次
- Does Writing with Language Models Reduce Content Diversity?Vishakh Padmakumar, He HeICLR 2024 · 被引用 173 次
- Working With AI to Persuade: Examining a Large Language Model's Ability to Generate Pro-Vaccination MessagesElise Karinshak, Sunny Xun Liu, Joon Sung Park, Jeffrey T. HancockCSCW 2023 · 被引用 163 次
相关 Paper
- Partnering with Generative AI: Experimental Evaluation of Model-Led and Human-Led Interaction in Human-AI Co-CreationSebastian Maier, Manuel Schneider, Stefan FeuerriegelCHI 2026 · 被引用 5 次
- An LLM Feature-based Framework for Dialogue Constructiveness AssessmentLexin Zhou, Youmna Farag, Andreas VlachosEMNLP 2024 · 被引用 3 次
- Supporting Human Raters with the Detection of Harmful Content Using Large Language ModelsKurt Thomas, Patrick Gage Kelley, David Tao, Sarah Meiklejohn 等S&P 2025
- LLM or Human? Perceptions of Trust and Quality in Research SummariesNil-Jana Akpinar, Sandeep Avula, Chia-Jung Lee, Brandon Dang 等CHI 2026 · 被引用 2 次
- Human Creativity in the Age of LLMs: Randomized Experiments on Divergent and Convergent ThinkingHarsh Kumar, Jonathan Vincentius, Ewan Jordan, Ashton AndersonCHI 2025 · 被引用 107 次
