Can Information Flows Suggest Targets for Interventions in Neural Circuits?
Praveen Venkatesh, Sanghamitra Dutta, Neil Ashim Mehta, Pulkit Grover
摘要
Motivated by neuroscientific and clinical applications, we empirically examine whether observational measures of information flow can suggest interventions. We do so by performing experiments on artificial neural networks in the context of fairness in machine learning, where the goal is to induce fairness in the system through interventions. Using our recently developed -information flow framework, we measure the flow of information about the true label (responsible for accuracy, and hence desirable), and separately, the flow of information about a protected attribute (responsible for bias, and hence undesirable) on the edges of a trained neural network. We then compare the flow magnitudes against the effect of intervening on those edges by pruning. We show that pruning edges that carry larger information flows about the protected attribute reduces bias at the output to a greater extent. This demonstrates that -information flow can meaningfully suggest targets for interventions, answering the title's question in the affirmative. We also evaluate bias-accuracy tradeoffs for different intervention strategies, to analyze how one might use estimates of desirable and undesirable information flows (here, accuracy and bias flows) to inform interventions that preserve the former while reducing the latter.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper4
- Algorithmic Transparency via Quantitative Input Influence: Theory and Experiments with Learning SystemsAnupam Datta, Shayak Sen, Yair ZickS&P 2016 · 被引用 774 次
- Is There a Trade-Off Between Fairness and Accuracy? A Perspective Using Mismatched Hypothesis TestingSanghamitra Dutta, Dennis Wei, Hazar Yueksel, Pin-Yu Chen 等ICML 2020 · 被引用 171 次
- An Information-Theoretic Quantification of Discrimination with Exempt FeaturesSanghamitra Dutta, Praveen Venkatesh, Piotr Mardziel, Anupam Datta 等AAAI 2020 · 被引用 34 次
- Bounding the fairness and accuracy of classifiers from population statisticsSivan Sabato, Elad Yom-TovICML 2020 · 被引用 18 次
相关 Paper
- The Fairness Hierarchy: A viewpoint from causal inferenceChengbo Zhang, Zhen Yao, Hao Pang, Changcheng LiICML 2026
- Towards Fairness-aware Adversarial Network PruningLei Zhang, Zhibo Wang, Xiaowei Dong, Yunhe Feng 等ICCV 2023 · 被引用 8 次
- Information-Theoretic Testing and Debugging of Fairness Defects in Deep Neural NetworksVerya Monjezi, Ashutosh Trivedi, Gang Tan, Saeid Tizpaz-NiariICSE 2023 · 被引用 47 次
- Understanding Fairness and Prediction Error through Subspace Decomposition and Influence AnalysisEnze Shi, Pankaj Bhagwat, Zhixian Yang, Linglong Kong 等NeurIPS 2025
- Adaptive fairness improvement based on causality analysisMengdi Zhang, Jun SunFSE 2022 · 被引用 35 次
