Algorithmic Collective Action in Machine Learning
Moritz Hardt, Eric Mazumdar, Celestine Mendler-Dünner, Tijana Zrnic
Abstract
We initiate a principled study of algorithmic collective action on digital platforms that deploy machine learning algorithms. We propose a simple theoretical model of a collective interacting with a firm's learning algorithm. The collective pools the data of participating individuals and executes an algorithmic strategy by instructing participants how to modify their own data to achieve a collective goal. We investigate the consequences of this model in three fundamental learning-theoretic settings: the case of a nonparametric optimal learning algorithm, a parametric risk minimizer, and gradient-based optimization. In each setting, we come up with coordinated algorithmic strategies and characterize natural success criteria as a function of the collective's size. Complementing our theory, we conduct systematic experiments on a skill classification task involving tens of thousands of resumes from a gig platform for freelancers. Through more than two thousand model training runs of a BERT-like language model, we see a striking correspondence emerge between our empirical observations and the predictions made by our theory. Taken together, our theory and experiments broadly support the conclusion that algorithmic collectives of exceedingly small fractional size can exert significant control over a platform's learning algorithm.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bf8f0fdd-2f4b-44c7-aa38-348320906235Cited by top-tier papers13
- Fine-Tuning Games: Bargaining and Adaptation for General-Purpose ModelsBenjamin Laufer, Jon M. Kleinberg, Hoda HeidariWWW 2024 · 27 citations
- Algorithmic Collective Action in Recommender Systems: Promoting Songs by Reordering PlaylistsJoachim Baumann, Celestine Mendler-DünnerNeurIPS 2024 · 12 citations
- The Role of Learning Algorithms in Collective ActionOmri Ben-Dov, Jake Fawkes, Samira Samadi, Amartya SanyalICML 2024 · 10 citations
- Prediction without Preclusion: Recourse Verification with Reachable SetsAvni Kothari, Bogdan Kulynych, Tsui-Wei Weng, Berk UstunICLR 2024 · 7 citations
- Look-Ahead Reasoning on Learning PlatformsHaiqing Zhu, Tijana Zrnic, Celestine Mendler-DünnerNeurIPS 2025 · 5 citations
Builds on9
- Witches' Brew: Industrial Scale Data Poisoning via Gradient MatchingJonas Geiping, Liam H. Fowl, W. Ronny Huang, Wojciech Czaja et al.ICLR 2021 · 268 citations
- Untargeted Backdoor Watermark: Towards Harmless and Stealthy Dataset Copyright ProtectionYiming Li, Yang Bai, Yong Jiang, Yong Yang et al.NeurIPS 2022 · 161 citations
- Who Leads and Who Follows in Strategic Classification?Tijana Zrnic, Eric Mazumdar, S. Shankar Sastry, Michael I. JordanNeurIPS 2021 · 76 citations
- Model-Targeted Poisoning Attacks with Provable ConvergenceFnu Suya, Saeed Mahloujifar, Anshuman Suri, David Evans et al.ICML 2021 · 52 citations
- LowKey: Leveraging Adversarial Attacks to Protect Social Media Users from Facial RecognitionValeriia Cherepanova, Micah Goldblum, Harrison Foley, Shiyuan Duan et al.ICLR 2021 · 52 citations
Related papers
- Statistical Collusion by Collectives on Learning PlatformsEtienne Gauthier, Francis Bach, Michael I. JordanICML 2025
- Stochastic Wage Suppression on Gig Platforms and How to Organize Against ItAna-Andreea Stoica, Celestine Mendler-Dünner, Moritz HardtWWW 2026
- Performative PowerMoritz Hardt, Meena Jagadeesan, Celestine Mendler-DünnerNeurIPS 2022 · 5 citations
- Gradient-Based Algorithms for Machine TeachingPei Wang, Kabir Nagrecha, Nuno VasconcelosCVPR 2021
- Reputation Agent: Prompting Fair Reviews in Gig MarketsCarlos Toxtli, Angela Richmond-Fuller, Saiph SavageWWW 2020 · 53 citations
