Algorithmic Collective Action in Machine Learning
Moritz Hardt, Eric Mazumdar, Celestine Mendler-Dünner, Tijana Zrnic
摘要
We initiate a principled study of algorithmic collective action on digital platforms that deploy machine learning algorithms. We propose a simple theoretical model of a collective interacting with a firm's learning algorithm. The collective pools the data of participating individuals and executes an algorithmic strategy by instructing participants how to modify their own data to achieve a collective goal. We investigate the consequences of this model in three fundamental learning-theoretic settings: the case of a nonparametric optimal learning algorithm, a parametric risk minimizer, and gradient-based optimization. In each setting, we come up with coordinated algorithmic strategies and characterize natural success criteria as a function of the collective's size. Complementing our theory, we conduct systematic experiments on a skill classification task involving tens of thousands of resumes from a gig platform for freelancers. Through more than two thousand model training runs of a BERT-like language model, we see a striking correspondence emerge between our empirical observations and the predictions made by our theory. Taken together, our theory and experiments broadly support the conclusion that algorithmic collectives of exceedingly small fractional size can exert significant control over a platform's learning algorithm.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Fine-Tuning Games: Bargaining and Adaptation for General-Purpose ModelsBenjamin Laufer, Jon M. Kleinberg, Hoda HeidariWWW 2024 · 被引用 27 次
- Algorithmic Collective Action in Recommender Systems: Promoting Songs by Reordering PlaylistsJoachim Baumann, Celestine Mendler-DünnerNeurIPS 2024 · 被引用 12 次
- The Role of Learning Algorithms in Collective ActionOmri Ben-Dov, Jake Fawkes, Samira Samadi, Amartya SanyalICML 2024 · 被引用 10 次
- Prediction without Preclusion: Recourse Verification with Reachable SetsAvni Kothari, Bogdan Kulynych, Tsui-Wei Weng, Berk UstunICLR 2024 · 被引用 7 次
- Look-Ahead Reasoning on Learning PlatformsHaiqing Zhu, Tijana Zrnic, Celestine Mendler-DünnerNeurIPS 2025 · 被引用 5 次
它引用的顶会 Paper9
- Witches' Brew: Industrial Scale Data Poisoning via Gradient MatchingJonas Geiping, Liam H. Fowl, W. Ronny Huang, Wojciech Czaja 等ICLR 2021 · 被引用 268 次
- Untargeted Backdoor Watermark: Towards Harmless and Stealthy Dataset Copyright ProtectionYiming Li, Yang Bai, Yong Jiang, Yong Yang 等NeurIPS 2022 · 被引用 161 次
- Who Leads and Who Follows in Strategic Classification?Tijana Zrnic, Eric Mazumdar, S. Shankar Sastry, Michael I. JordanNeurIPS 2021 · 被引用 76 次
- Model-Targeted Poisoning Attacks with Provable ConvergenceFnu Suya, Saeed Mahloujifar, Anshuman Suri, David Evans 等ICML 2021 · 被引用 52 次
- LowKey: Leveraging Adversarial Attacks to Protect Social Media Users from Facial RecognitionValeriia Cherepanova, Micah Goldblum, Harrison Foley, Shiyuan Duan 等ICLR 2021 · 被引用 52 次
相关 Paper
- Statistical Collusion by Collectives on Learning PlatformsEtienne Gauthier, Francis Bach, Michael I. JordanICML 2025
- Stochastic Wage Suppression on Gig Platforms and How to Organize Against ItAna-Andreea Stoica, Celestine Mendler-Dünner, Moritz HardtWWW 2026
- Performative PowerMoritz Hardt, Meena Jagadeesan, Celestine Mendler-DünnerNeurIPS 2022 · 被引用 5 次
- Gradient-Based Algorithms for Machine TeachingPei Wang, Kabir Nagrecha, Nuno VasconcelosCVPR 2021
- Reputation Agent: Prompting Fair Reviews in Gig MarketsCarlos Toxtli, Angela Richmond-Fuller, Saiph SavageWWW 2020 · 被引用 53 次
