Bias in Motion: Theoretical Insights into the Dynamics of Bias in SGD Training
Anchit Jain, Rozhin Nobahari, Aristide Baratin, Stefano Sarao Mannelli
摘要
Machine learning systems often acquire biases by leveraging undesired features in the data, impacting accuracy variably across different sub-populations. Current understanding of bias formation mostly focuses on the initial and final stages of learning, leaving a gap in knowledge regarding the transient dynamics. To address this gap, this paper explores the evolution of bias in a teacher-student setup modeling different data sub-populations with a Gaussian-mixture model. We provide an analytical description of the stochastic gradient descent dynamics of a linear classifier in this setting, which we prove to be exact in high dimension. Notably, our analysis reveals how different properties of sub-populations influence bias at different timescales, showing a shifting preference of the classifier during training. Applying our findings to fairness and robustness, we delineate how and when heterogeneous data and spurious features can generate and amplify bias. We empirically validate our results in more complex scenarios by training deeper networks on synthetic and real datasets, including CIFAR10, MNIST, and CelebA.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Reward Models Inherit Value Biases from PretrainingBrian R. Christian, Jessica A. F. Thompson, Elle Michelle Yang, Vincent Adam 等ICLR 2026 · 被引用 4 次
- Sharp description of local minima in the loss landscape of high-dimensional two-layer ReLU neural networksJie Huang, Bruno Loureiro, Stefano Sarao MannelliICML 2026 · 被引用 1 次
- Do We Always Need the Simplicity Bias? Looking for Optimal Inductive Biases in the WildDamien Teney, Liangze Jiang, Florin Gogianu, Ehsan AbbasnejadCVPR 2025
- A Theory of Initialisation's Impact on SpecialisationDevon Jarvis, Sebastian Lee, Clémentine Carla Juliette Dominé, Andrew M. Saxe 等ICLR 2025
- Optimal Protocols for Continual Learning via Statistical Physics and Control TheoryFrancesco Mori, Stefano Sarao Mannelli, Francesca MignaccoICLR 2025
它引用的顶会 Paper11
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 被引用 1,578 次
- Just Train Twice: Improving Group Robustness without Training Group InformationEvan Zheran Liu, Behzad Haghgoo, Annie S. Chen, Aditi Raghunathan 等ICML 2021 · 被引用 683 次
- The Pitfalls of Simplicity Bias in Neural NetworksHarshay Shah, Kaustav Tamuly, Aditi Raghunathan, Prateek Jain 等NeurIPS 2020 · 被引用 503 次
- An Investigation of Why Overparameterization Exacerbates Spurious CorrelationsShiori Sagawa, Aditi Raghunathan, Pang Wei Koh, Percy LiangICML 2020 · 被引用 436 次
- Understanding the failure modes of out-of-distribution generalizationVaishnavh Nagarajan, Anders Andreassen, Behnam NeyshaburICLR 2021 · 被引用 205 次
相关 Paper
- The Implicit Bias of Heterogeneity towards Invariance: A Study of Multi-Environment Matrix SensingYang Xu, Yihong Gu, Cong FangNeurIPS 2024 · 被引用 1 次
- Stochastic Collapse: How Gradient Noise Attracts SGD Dynamics Towards Simpler SubnetworksFeng Chen, Daniel Kunin, Atsushi Yamamura, Surya GanguliNeurIPS 2023 · 被引用 52 次
- Understanding the Impact of Adversarial Robustness on Accuracy DisparityYuzheng Hu, Fan Wu, Hongyang Zhang, Han ZhaoICML 2023 · 被引用 11 次
- An Effective Theory of Bias AmplificationArjun Subramonian, Samuel J. Bell, Levent Sagun, Elvis DohmatobICLR 2025
- Fairness in Forecasting and Learning Linear Dynamical SystemsQuan Zhou, Jakub Marecek, Robert N. ShortenAAAI 2021 · 被引用 10 次
