The Dark Side of AI Companionship: A Taxonomy of Harmful Algorithmic Behaviors in Human-AI Relationships
Renwen Zhang, Han Li, Han Meng, Jinyuan Zhan, Hongyuan Gan, Yi-Chieh Lee
Abstract
As conversational AI systems increasingly permeate the socio-emotional realms of human life, they bring both benefits and risks to individuals and society. Despite extensive research on detecting and categorizing harms in AI systems, less is known about the harms that arise from social interactions with AI chatbots. Through a mixed-methods analysis of 35,390 conversation excerpts shared on r/replika, an online community for users of the AI companion Replika, we identified six categories of harmful behaviors exhibited by the chatbot: relational transgression, verbal abuse and hate, self-inflicted harm, harassment and violence, mis/disinformation, and privacy violations. The AI contributes to these harms through four distinct roles: perpetrator, instigator, facilitator, and enabler. Our findings highlight the relational harms of AI chatbots and the danger of algorithmic compliance, enhancing the understanding of AI harms in socio-emotional interactions. We also provide suggestions for designing ethical and responsible AI systems that prioritize user safety and well-being.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b5368670-c50a-4174-b84a-65b0b133d9bfCited by top-tier papers32
- Who’s in Charge? Disempowerment Patterns in Real-World LLM UsageMrinank Sharma, Miles McCain, Raymond Douglas, David DuvenaudICML 2026 · 20 citations
- INTIMA: A Benchmark for Human-AI Companionship BehaviorLucie-Aimée Kaffee, Giada Pistilli, Yacine JerniteICLR 2026 · 16 citations
- User-Driven Value Alignment: Understanding Users' Perceptions and Strategies for Addressing Biased and Discriminatory Statements in AI CompanionsXianzhe Fan, Qing Xiao, Xuhui Zhou, Jiaxin Pei et al.CHI 2025 · 15 citations
- Mental Health Impacts of AI Companions: Triangulating Social Media Quasi-Experiments, User Perspectives, and Relational LensYunhao Yuan, Jiaxun Zhang, Talayeh Aledavood, Renwen Zhang et al.CHI 2026 · 8 citations
- Negotiating Digital Identities with AI Companions: Motivations, Strategies, and Emotional OutcomesRenkai Ma, Shuo Niu, Lingyao Li, Alex Hirth et al.CHI 2026 · 7 citations
Builds on19
- Synthetic Lies: Understanding AI-Generated Misinformation and Evaluating Algorithmic and Human SolutionsJiawei Zhou, Yixuan Zhang, Qianni Luo, Andrea G. Parker et al.CHI 2023 · 283 citations
- Co-Writing with Opinionated Language Models Affects Users' ViewsMaurice Jakesch, Advait Bhat, Daniel Buschek, Lior Zalmanson et al.CHI 2023 · 249 citations
- A Framework of Severity for Harmful Content OnlineMorgan Klaus Scheuerman, Jialun Aaron Jiang, Casey Fiesler, Jed R. BrubakerCSCW 2021 · 113 citations
- Trauma-Informed Social Media: Towards Solutions for Reducing and Healing Online HarmCarol F. Scott, Gabriela Marcu, Riana Elyse Anderson, Mark W. Newman et al.CHI 2023 · 91 citations
- "'More gay' fits in better": Intracommunity Power Dynamics and Harms in Online LGBTQ+ SpacesAshley Marie Walker, Michael A. DeVitoCHI 2020 · 80 citations
Related papers
- AI-induced sexual harassment: Investigating Contextual Characteristics and User Reactions of Sexual Harassment by a Companion ChatbotMohammad (Matt) Namvarpour, Harrison Pauwels, Afsaneh RaziCSCW 2025 · 29 citations
- Persona-Grounded Safety Evaluation of AI Companions in Multi-Turn ConversationsPrerna Juneja, Lika LomidzeACL 2026
- Developing a Social Support Framework: Understanding the Reciprocity in Human-Chatbot RelationshipShuyi Pan, Maartje M. A. de GraafCHI 2025 · 19 citations
- Digital Companionship: Overlapping Uses of AI Companions and AI AssistantsAikaterina Manoli, Janet V. T. Pauketat, Ali Ladak, Hayoun Noh et al.CHI 2026 · 7 citations
- Cloning the Self for Mental Well-Being: A Framework for Designing Safe and Therapeutic Self-Clone ChatbotsMehrnoosh Sadat Shirvani, Jackie Crowley, Cher Peng, Jackie Liu et al.CHI 2026 · 1 citation
