AIM: Attributing, Interpreting, Mitigating Data Unfairness
Zhining Liu, Ruizhong Qiu, Zhichen Zeng, Yada Zhu, Hendrik F. Hamann, Hanghang Tong
Abstract
Data collected in the real world often encapsulates historical discrimination against disadvantaged groups and individuals. Existing fair machine learning (FairML) research has predominantly focused on mitigating discriminative bias in the model prediction, with far less effort dedicated towards exploring how to trace biases present in the data, despite its importance for the transparency and interpretability of FairML. To fill this gap, we investigate a novel research problem: discovering samples that reflect biases/prejudices from the training data. Grounding on the existing fairness notions, we lay out a sample bias criterion and propose practical algorithms for measuring and countering sample bias. The derived bias score provides intuitive sample-level attribution and explanation of historical bias in data. On this basis, we further design two FairML strategies via sample-bias-informed minimal data editing. They can mitigate both group and individual unfairness at the cost of minimal or zero predictive utility loss. Extensive experiments and analyses on multiple real-world datasets demonstrate the effectiveness of our methods in explaining and mitigating unfairness. Code is available at https://github.com/ZhiningLiu1998/AIM.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Joint Optimal Transport and Embedding for Network AlignmentQi Yu, Zhichen Zeng, Yuchen Yan, Lei Ying et al.WWW 2025 · 17 citations
- Gradient Compressed Sensing: A Query-Efficient Gradient Estimator for High-Dimensional Zeroth-Order OptimizationRuizhong Qiu, Hanghang TongICML 2024 · 12 citations
- Graph homophily booster: Reimagining the role of discrete features in heterophilic graph learningRuizhong Qiu, Ting-Wei Li, Gaotang Li, Hanghang TongICLR 2026 · 2 citations
- APEX2: Adaptive and Extreme Summarization for Personalized Knowledge GraphsZihao Li, Dongqi Fu, Mengting Ai, Jingrui HeKDD 2025 · 1 citation
- Ask, and it shall be given: On the Turing completeness of promptingRuizhong Qiu, Zhe Xu, Wenxuan Bao, Hanghang TongICLR 2025
Builds on15
- Self-paced Ensemble for Highly Imbalanced Massive Data ClassificationZhining Liu, Wei Cao, Zhifeng Gao, Jiang Bian et al.ICDE 2020 · 172 citations
- Robust Optimization for Fairness with Noisy Protected GroupsSerena Lutong Wang, Wenshuo Guo, Harikrishna Narasimhan, Andrew Cotter et al.NeurIPS 2020 · 134 citations
- Training individually fair ML models with sensitive subspace robustnessMikhail Yurochkin, Amanda Bower, Yuekai SunICLR 2020 · 123 citations
- Two Simple Ways to Learn Individual Fairness Metrics from DataDebarghya Mukherjee, Mikhail Yurochkin, Moulinath Banerjee, Yuekai SunICML 2020 · 109 citations
- Dynamic Knowledge Graph AlignmentYuchen Yan, Lihui Liu, Yikun Ban, Baoyu Jing et al.AAAI 2021 · 100 citations
Related papers
- Fairway: a way to build fair ML softwareJoymallya Chakraborty, Suvodeep Majumder, Zhe Yu, Tim MenziesFSE 2020 · 131 citations
- HIFI: Explaining and Mitigating Algorithmic Bias Through the Lens of Game-Theoretic InteractionsLingfeng Zhang, Zhaohui Wang, Yueling Zhang, Min Zhang et al.ICSE 2025 · 2 citations
- Software Fairness Dilemma: Is Bias Mitigation a Zero-Sum Game?Zhenpeng Chen, Xinyue Li, Jie M. Zhang, Weisong Sun et al.FSE 2025
- Social Bias Meets Data Bias: The Impacts of Labeling and Measurement Errors on Fairness CriteriaYiqiao Liao, Parinaz NaghizadehAAAI 2023 · 15 citations
- Fairness and Explainability: Bridging the Gap towards Fair Model ExplanationsYuying Zhao, Yu Wang, Tyler DerrAAAI 2023 · 27 citations
