Long-Term Fairness with Unknown Dynamics
Tongxin Yin, Reilly Raab, Mingyan Liu, Yang Liu
Abstract
While machine learning can myopically reinforce social inequalities, it may also be used to dynamically seek equitable outcomes. In this paper, we formalize long-term fairness in the context of online reinforcement learning. This formulation can accommodate dynamical control objectives, such as driving equity inherent in the state of a population, that cannot be incorporated into static formulations of fairness. We demonstrate that this framing allows an algorithm to adapt to unknown dynamics by sacrificing short-term incentives to drive a classifier-population system towards more desirable equilibria. For the proposed setting, we develop an algorithm that adapts recent work in online learning. We prove that this algorithm achieves simultaneous probabilistic bounds on cumulative loss and cumulative violations of fairness (as statistical regularities between demographic groups). We compare our proposed algorithm to the repeated retraining of myopic classifiers, as a baseline, and to a deep reinforcement learning algorithm that lacks safety guarantees. Our experiments model human populations according to evolutionary game theory and integrate real-world datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Adapting Static Fairness to Sequential Decision-Making: Bias Mitigation Strategies towards Equal Long-term Benefit RateYuancheng Xu, Chenghao Deng, Yanchao Sun, Ruijie Zheng et al.ICML 2024 · 7 citations
- FairSense: Long-Term Fairness Analysis of ML-Enabled SystemsYining She, Sumon Biswas, Christian Kästner, Eunsuk KangICSE 2025 · 4 citations
- Enhancing Group Fairness in Online Settings Using Oblique Decision ForestsSomnath Basu Roy Chowdhury, Nicholas Monath, Ahmad Beirami, Rahul Kidambi et al.ICLR 2024 · 3 citations
- Retention Depolarization in Recommender SystemXiaoying Zhang, Hongning Wang, Yang LiuWWW 2024 · 2 citations
- Fair Participation via Sequential PoliciesReilly Raab, Ross Boczar, Maryam Fazel, Yang LiuAAAI 2024 · 1 citation
Builds on12
- Performative PredictionJuan C. Perdomo, Tijana Zrnic, Celestine Mendler-Dünner, Moritz HardtICML 2020 · 422 citations
- Natural Policy Gradient Primal-Dual Method for Constrained Markov Decision ProcessesDongsheng Ding, Kaiqing Zhang, Tamer Basar, Mihailo R. JovanovicNeurIPS 2020 · 252 citations
- CRPO: A New Approach for Safe Reinforcement Learning with Convergence GuaranteeTengyu Xu, Yingbin Liang, Guanghui LanICML 2021 · 171 citations
- Learning Policies with Zero or Bounded Constraint Violation for Constrained MDPsTao Liu, Ruida Zhou, Dileep Kalathil, Panganamala R. Kumar et al.NeurIPS 2021 · 110 citations
- Achieving Zero Constraint Violation for Constrained Reinforcement Learning via Primal-Dual ApproachQinbo Bai, Amrit Singh Bedi, Mridul Agarwal, Alec Koppel et al.AAAI 2022 · 69 citations
Related papers
- Automating Data Annotation under Strategic Human Agents: Risks and Potential SolutionsTian Xie, Xueru ZhangNeurIPS 2024 · 12 citations
- Causal Modeling for Fairness In Dynamical SystemsElliot Creager, David Madras, Toniann Pitassi, Richard S. ZemelICML 2020 · 72 citations
- How do fair decisions fare in long-term qualification?Xueru Zhang, Ruibo Tu, Yang Liu, Mingyan Liu et al.NeurIPS 2020 · 87 citations
- Adaptive Fairness-Aware Online Meta-Learning for Changing EnvironmentsChen Zhao, Feng Mi, Xintao Wu, Kai Jiang et al.KDD 2022 · 20 citations
- Fairness Transferability Subject to Bounded Distribution ShiftYatong Chen, Reilly Raab, Jialu Wang, Yang LiuNeurIPS 2022 · 40 citations
