It's Trying Too Hard To Look Real: Deepfake Moderation Mistakes and Identity-Based Bias
Jaron Mink, Miranda Wei, Collins W. Munyendo, Kurt Hugenberg, Tadayoshi Kohno, Elissa M. Redmiles, Gang Wang
摘要
Online platforms employ manual human moderation to distinguish human-created social media profiles from deepfake-generated ones. Biased misclassification of real profiles as artificial can harm general users as well as specific identity groups; however, no work has yet systematically investigated such mistakes and biases. We conducted a user study (𝑛=695) that investigates how 1) the identity of the profile, 2) whether the moderator shares that identity, and 3) components of a profile shown affect the perceived artificiality of the profile. We find statistically significant biases in people's moderation of LinkedIn profiles based on all three factors. Further, upon examining how moderators make decisions, we find they rely on mental models of AI and attackers, as well as typicality expectations (how they think the world works). The latter includes reliance on race/gender stereotypes. Based on our findings, we synthesize recommendations for the design of moderation interfaces, moderation teams, and security training.
• Security and privacy → Social aspects of security and privacy; • Human-centered computing → Empirical studies in HCI .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Characterizing Photorealism and Artifacts in Diffusion Model-Generated ImagesNegar Kamali, Karyn Nakamura, Aakriti Kumar, Angelos Chatzimparmpas 等CHI 2025 · 被引用 23 次
- Public Opinions About Copyright for AI-Generated Art: The Role of Egocentricity, Competition, and ExperienceGabriel Lima, Nina Grgic-Hlaca, Elissa M. RedmilesCHI 2025 · 被引用 18 次
- 'There Has To Be a Lot That We're Missing': Moderating AI-Generated Content on RedditTravis Lloyd, Joseph Reagle, Mor NaamanCSCW 2025 · 被引用 6 次
- Behind the Same Mask: Understanding the Practice of Spontaneous Collective Anonymity on Chinese Social PlatformsSuqi Lou, Weijun Li, Chao Zhang, Shi Chen 等CSCW 2025 · 被引用 4 次
- Governance of AI-Generated Content: A Case Study on Social Media PlatformsLan Gao, Abani Ahmed, Oscar Chen, Margaux Reyl 等CHI 2026 · 被引用 3 次
它引用的顶会 Paper8
- Disproportionate Removals and Differing Content Moderation Experiences for Conservative, Transgender, and Black Social Media Users: Marginalization and Moderation Gray AreasOliver L. Haimson, Daniel Delmonaco, Peipei Nie, Andrea WegnerCSCW 2021 · 被引用 287 次
- Self-supervised Learning of Adversarial Example: Towards Good Generalizations for Deepfake DetectionLiang Chen, Yong Zhang, Yibing Song, Lingqiao Liu 等CVPR 2022 · 被引用 251 次
- Conformity of Eating Disorders through Content ModerationJessica L. Feuston, Alex S. Taylor, Anne Marie PiperCSCW 2020 · 被引用 82 次
- Seeing is Believing: Exploring Perceptual Differences in DeepFake VideosRashid Tahir, Brishna Batool, Hira Jamshed, Mahnoor Jameel 等CHI 2021 · 被引用 78 次
- How Experts Detect Phishing Scam EmailsRick WashCSCW 2020 · 被引用 76 次
相关 Paper
- DeepPhish: Understanding User Trust Towards Artificially Generated Profiles in Online Social NetworksJaron Mink, Licheng Luo, Natã M. Barbosa, Olivia Figueira 等USENIX Security 2022
- "Better Be Computer or I'm Dumb": A Large-Scale Evaluation of Humans as Audio Deepfake DetectorsKevin Warren, Tyler Tucker, Anna Crowder, Daniel Olszewski 等CCS 2024 · 被引用 9 次
- Effect of AI Performance, Risk Perception, and Trust on Human Dependence in Deepfake Detection AI SystemYingfan Zhou, Ester Chen, Manasa Pisipati, Aiping Xiong 等CSCW 2025 · 被引用 2 次
- "It Matches My Worldview": Examining Perceptions and Attitudes Around Fake VideosFarhana Shahid, Srujana Kamath, Annie Sidotam, Vivian Jiang 等CHI 2022 · 被引用 38 次
- Labeling Synthetic Content: User Perceptions of Label Designs for AI-Generated Content on Social MediaDilrukshi Gamage, Dilki Sewwandi, Min Zhang, Arosha K. BandaraCHI 2025 · 被引用 25 次
