User-Driven Value Alignment: Understanding Users' Perceptions and Strategies for Addressing Biased and Discriminatory Statements in AI Companions
Xianzhe Fan, Qing Xiao, Xuhui Zhou, Jiaxin Pei, Maarten Sap, Zhicong Lu, Hong Shen
Abstract
Content Warning: This paper presents textual examples that may be offensive or upsetting.
Large language model-based AI companions are increasingly viewed by users as friends or romantic partners, leading to deep emotional bonds. However, they can generate biased, discriminatory, and harmful outputs. Recently, users are taking the initiative to address these harms and re-align AI companions. We introduce the concept of user-driven value alignment, where users actively identify, challenge, and attempt to correct AI outputs they perceive as harmful, aiming to guide the AI to better align with their values. We analyzed 77 social media posts about discriminatory AI statements and conducted semi-structured interviews with 20 experienced users. Our analysis revealed six common types of discriminatory statements perceived by users, how users make sense of those AI behaviors, and seven user-driven alignment strategies, such as gentle persuasion and anger expression. We discuss implications for supporting user-driven value alignment in future AI systems, where users and their communities have greater agency.
• Human-centered computing → Empirical studies in HCI .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6b55048e-d00a-44c9-a6c7-ee9e78ccf0cdCited by top-tier papers15
- Interaction Context Often Increases Sycophancy in LLMsShomik Jain, Charlotte Park, Matt Viana, Ashia Wilson et al.CHI 2026 · 12 citations
- Relational Dissonance in Human-AI Interactions: The Case of Knowledge WorkEmrecan Gulay, Eleonora Picco, Enrico Glerean, Corinna CoupetteCHI 2026 · 8 citations
- Negotiating Digital Identities with AI Companions: Motivations, Strategies, and Emotional OutcomesRenkai Ma, Shuo Niu, Lingyao Li, Alex Hirth et al.CHI 2026 · 7 citations
- AI and My Values: User Perceptions of LLMs' Ability to Extract, Embody, and Explain Human Values from Casual ConversationsBhada Yun, Renn Su, April Yi WangCHI 2026 · 7 citations
- POET: Supporting Prompting Creativity and Personalization with Automated Expansion of Text-to-Image GenerationEvans Xu Han, Alice Qian Zhang, Haiyi Zhu, Hong Shen et al.UIST 2025 · 5 citations
Builds on27
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- LIMA: Less Is More for AlignmentChunting Zhou, Pengfei Liu, Puxin Xu, Srinivasan Iyer et al.NeurIPS 2023 · 1,486 citations
- A Survey on In-context LearningQingxiu Dong, Lei Li, Damai Dai, Ce Zheng et al.EMNLP 2024 · 479 citations
- Principle-Driven Self-Alignment of Language Models from Scratch with Minimal Human SupervisionZhiqing Sun, Yikang Shen, Qinhong Zhou, Hongxin Zhang et al.NeurIPS 2023 · 463 citations
Related papers
- Unintended Harms of Value-Aligned LLMs: Psychological and Empirical InsightsSooyung Choi, Jaehyeok Lee, Xiaoyuan Yi, Jing Yao et al.ACL 2025
- The Typing Cure: Experiences with Large Language Model Chatbots for Mental Health SupportInhwa Song, Sachin R. Pendse, Neha Kumar, Munmun De ChoudhuryCSCW 2025 · 41 citations
- Caught in a Mafia Romance: How Users Explore Intimate Narratives with ChatbotsJulia B. Kieserman, Cat Mai, Sara Lignell, Lucy Qin et al.CHI 2026 · 1 citation
- Toward User-Driven Algorithm Auditing: Investigating users' strategies for uncovering harmful algorithmic behaviorAlicia DeVos, Aditi Dhabalia, Hong Shen, Kenneth Holstein et al.CHI 2022 · 96 citations
- "Please, don't kill the only model that still feels human": Understanding the #Keep4o BacklashHuiqian LaiCHI 2026 · 5 citations
