Data Minimization at Inference Time
Cuong Tran, Ferdinando Fioretto
摘要
In domains with high stakes such as law, recruitment, and healthcare, learning models frequently rely on sensitive user data for inference, necessitating the complete set of features. This not only poses significant privacy risks for individuals but also demands substantial human effort from organizations to verify information accuracy. This paper asks whether it is necessary to use all input features for accurate predictions at inference time. The paper demonstrates that, in a personalized setting, individuals may only need to disclose a small subset of their features without compromising decision-making accuracy. The paper also provides an efficient sequential algorithm to determine the appropriate attributes for each individual to provide. Evaluations across various learning tasks show that individuals can potentially report as little as 10% of their information while maintaining the same accuracy level as a model that employs the full set of user information.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper4
- Deep Learning with Differential PrivacyMartín Abadi, Andy Chu, Ian J. Goodfellow, H. Brendan McMahan 等CCS 2016 · 被引用 7,620 次
- Improving model calibration with accuracy versus uncertainty optimizationRanganath Krishnan, Omesh TickooNeurIPS 2020 · 被引用 217 次
- Robust and differentially private mean estimationXiyang Liu, Weihao Kong, Sham M. Kakade, Sewoong OhNeurIPS 2021 · 被引用 87 次
- Auditing Black-Box Prediction Models for Data Minimization ComplianceBashir Rastegarpanah, Krishna P. Gummadi, Mark CrovellaNeurIPS 2021 · 被引用 24 次
相关 Paper
- DISCO: Dynamic and Invariant Sensitive Channel Obfuscation for Deep Neural NetworksAbhishek Singh, Ayush Chopra, Ethan Garza, Emily Zhang 等CVPR 2021
- When Machine Learning Gets Personal: Evaluating Prediction and ExplanationLouisa Cornelis, Guillermo Bernardez, Haewon Jeong, Nina MiolaneICLR 2026
- Fair Learning with Private Demographic DataHussein Mozannar, Mesrob I. Ohannessian, Nathan SrebroICML 2020 · 被引用 85 次
- Participatory Personalization in ClassificationHailey Joren, Chirag Nagpal, Katherine A. Heller, Berk UstunNeurIPS 2023 · 被引用 7 次
- Learning Fair Naive Bayes Classifiers by Discovering and Eliminating Discrimination PatternsYooJung Choi, Golnoosh Farnadi, Behrouz Babaki, Guy Van den BroeckAAAI 2020 · 被引用 31 次
