Effects of Algorithmic Flagging on Fairness: Quasi-experimental Evidence from Wikipedia
Nathan TeBlunthuis, Benjamin Mako Hill, Aaron Halfaker
Abstract
Online community moderators often rely on social signals such as whether or not a user has an account or a profile page as clues that users may cause problems. Reliance on these clues can lead to "overprofiling'' bias when moderators focus on these signals but overlook the misbehavior of others. We propose that algorithmic flagging systems deployed to improve the efficiency of moderation work can also make moderation actions more fair to these users by reducing reliance on social signals and making norm violations by everyone else more visible. We analyze moderator behavior in Wikipedia as mediated by RCFilters, a system which displays social signals and algorithmic flags, and estimate the causal effect of being flagged on moderator actions. We show that algorithmically flagged edits are reverted more often, especially those by established editors with positive social signals, and that flagging decreases the likelihood that moderation actions will be undone. Our results suggest that algorithmic flagging systems can lead to increased fairness in some contexts but that the relationship is complex and contingent.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 53a3e8c6-9743-490e-bd24-786ad1b59f8fCited by top-tier papers6
- Decolonizing Content Moderation: Does Uniform Global Community Standard Resemble Utopian Equality or Western Power Hegemony?Farhana Shahid, Aditya VashisthaCHI 2023 · 63 citations
- Automated Content Moderation Increases Adherence to Community GuidelinesManoel Horta Ribeiro, Justin Cheng, Robert WestWWW 2023 · 53 citations
- Taboo and Collaborative Knowledge Production: Evidence from WikipediaKaylea Champion, Benjamin Mako HillCSCW 2023 · 2 citations
- Governing Together: Toward Infrastructure for Community-Run Social MediaSohyeon Hwang, Sophie Rollins, Thatiany Andrade Nunes, Yuhan Liu et al.CHI 2026 · 1 citation
- The Relational Origins of Rules in Online CommunitiesCharles Kiene, Sohyeon Hwang, Carl Colglazier, Nathan TeBlunthuis et al.CSCW 2026 · 1 citation
Builds on1
Related papers
- The Risks, Benefits, and Consequences of Prepublication Moderation: Evidence from 17 Wikipedia Language EditionsChau Tran, Kaylea Champion, Benjamin Mako Hill, Rachel GreenstadtCSCW 2022 · 4 citations
- Challenges in Restructuring Community-based ModerationChau Tran, Kejsi Take, Kaylea Champion, Benjamin Mako Hill et al.CSCW 2024 · 1 citation
- Keeping Community in the Loop: Understanding Wikipedia Stakeholder Values for Machine Learning-Based SystemsC. Estelle Smith, Bowen Yu, Anjali Srivastava, Aaron Halfaker et al.CHI 2020 · 77 citations
- The Use of Negative Interface Cues to Change Perceptions of Online Retributive HarassmentSong Mi Lee, Andrea K. Thomer, Cliff LampeCSCW 2022 · 9 citations
- Proactive Moderation of Online Discussions: Existing Practices and the Potential for Algorithmic SupportCharlotte Schluger, Jonathan P. Chang, Cristian Danescu-Niculescu-Mizil, Karen LevyCSCW 2022 · 41 citations
