Generalizing Orthogonalization for Models with Non-Linearities
David Rügamer, Chris Kolb, Tobias Weber, Lucas Kook, Thomas Nagler
摘要
The complexity of black-box algorithms can lead to various challenges, including the introduction of biases. These biases present immediate risks in the algorithms' application. It was, for instance, shown that neural networks can deduce racial information solely from a patient's X-ray scan, a task beyond the capability of medical experts. If this fact is not known to the medical expert, automatic decision-making based on this algorithm could lead to prescribing a treatment (purely) based on racial information. While current methodologies allow for the "orthogonalization" or "normalization" of neural networks with respect to such information, existing approaches are grounded in linear models. Our paper advances the discourse by introducing corrections for non-linearities such as ReLU activations. Our approach also encompasses scalar and tensorvalued predictions, facilitating its integration into neural network architectures. Through extensive experiments, we validate our method's effectiveness in safeguarding sensitive data in generalized linear models, normalizing convolutional neural networks for metadata, and rectifying pre-existing embeddings for undesired attributes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Fair Infinitesimal Jackknife: Mitigating the Influence of Biased Training Data Points Without RefittingPrasanna Sattigeri, Soumya Ghosh, Inkit Padhi, Pierre L. Dognin 等NeurIPS 2022 · 被引用 36 次
- Controlling Directions Orthogonal to a ClassifierYilun Xu, Hao He, Tianxiao Shen, Tommi S. JaakkolaICLR 2022 · 被引用 20 次
- Fair Generalized Linear Models with a Convex PenaltyHyungrok Do, Preston Putzel, Axel S. Martin, Padhraic Smyth 等ICML 2022 · 被引用 19 次
- A New PHO-rmula for Improved Performance of Semi-Structured NetworksDavid RügamerICML 2023 · 被引用 11 次
- Scalable Infomin LearningYanzhi Chen, Weihao Sun, Yingzhen Li, Adrian WellerNeurIPS 2022 · 被引用 10 次
相关 Paper
- Disparate Impact on Group Accuracy of Linearization for Private InferenceSaswat Das, Marco Romanelli, Ferdinando FiorettoICML 2024 · 被引用 4 次
- Divergence-Free Neural Networks with Application to Image DenoisingSébastien Herbreteau, Etienne MeunierICLR 2026
- Normalization-Equivariant Neural Networks with Application to Image DenoisingSébastien Herbreteau, Emmanuel Moebel, Charles KervrannNeurIPS 2023 · 被引用 20 次
- Sound and Complete Neural Network Repair with Minimality and Locality GuaranteesFeisi Fu, Wenchao LiICLR 2022 · 被引用 36 次
- Discover the Unknown Biased Attribute of an Image ClassifierZhiheng Li, Chenliang XuICCV 2021 · 被引用 55 次
