Privacy-Preserving Feature Selection with Secure Multiparty Computation
Xiling Li, Rafael Dowsley, Martine De Cock
Abstract
Existing work on privacy-preserving machine learning with Secure Multiparty Computation (MPC) is almost exclusively focused on model training and on inference with trained models, thereby overlooking the important data pre-processing stage. In this work, we propose the first MPC based protocol for private feature selection based on the filter method, which is independent of model training, and can be used in combination with any MPC protocol to rank features. We propose an efficient feature scoring protocol based on Gini impurity to this end. To demonstrate the feasibility of our approach for practical data science, we perform experiments with the proposed MPC protocols for feature selection in a commonly used machine-learning-as-a-service configuration where computations are outsourced to multiple servers, with semi-honest and with malicious adversaries. Regarding effectiveness, we show that secure feature selection with the proposed protocols improves the accuracy of classifiers on a variety of real-world data sets, without leaking information about the feature values or even which features were selected. Regarding efficiency, we document runtimes ranging from several seconds to an hour for our protocols to finish, depending on the size of the data set and the security settings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1b1f87d2-b3a0-4a8d-a667-d672637cce7fCited by top-tier papers6
- FedSDG-FS: Efficient and Secure Feature Selection for Vertical Federated LearningAnran Li, Hongyi Peng, Lan Zhang, Jiahui Huang et al.INFOCOM 2023 · 50 citations
- On the Gini-impurity Preservation For Privacy Random ForestsXinran Xie, Man-Jie Yuan, Xuetong Bai, Wei Gao et al.NeurIPS 2023 · 17 citations
- FEAST: A Communication-efficient Federated Feature Selection Framework for Relational DataRui Fu, Yuncheng Wu, Quanqing Xu, Meihui ZhangSIGMOD 2023 · 17 citations
- Ents: An Efficient Three-party Training Framework for Decision Trees by Communication OptimizationGuopeng Lin, Weili Han, Wenqiang Ruan, Ruisheng Zhou et al.CCS 2024 · 3 citations
- Comet: Accelerating Private Inference for Large Language Model by Predicting Activation SparsityGuang Yan, Yuhui Zhang, Zimu Guo, Lutan Zhao et al.S&P 2025
Builds on6
- SecureML: A System for Scalable Privacy-Preserving Machine LearningPayman Mohassel, Yupeng ZhangS&P 2017 · 2,107 citations
- High-Throughput Semi-Honest Secure Three-Party Computation with an Honest MajorityToshinori Araki, Jun Furukawa, Yehuda Lindell, Ariel Nof et al.CCS 2016 · 463 citations
- CrypTFlow: Secure TensorFlow InferenceNishant Kumar, Mayank Rathee, Nishanth Chandran, Divya Gupta et al.S&P 2020 · 276 citations
- QUOTIENT: Two-Party Secure Neural Network Training and PredictionNitin Agrawal, Ali Shahin Shamsabadi, Matt J. Kusner, Adrià GascónCCS 2019 · 241 citations
- Fantastic Four: Honest-Majority Four-Party Secure Computation With Malicious SecurityAnders P. K. Dalskov, Daniel Escudero, Marcel KellerUSENIX Security 2021 · 174 citations
Related papers
- ABNN2: secure two-party arbitrary-bitwidth quantized neural network predictionsLiyan Shen, Ye Dong, Binxing Fang, Jinqiao Shi et al.DAC 2022 · 11 citations
- Differentially Private Selection from Secure Distributed ComputingIvan Damgård, Hannah Keller, Boel Nelson, Claudio Orlandi et al.WWW 2024 · 3 citations
- Privacy-Preserving Video Classification with Convolutional Neural NetworksSikha Pentyala, Rafael Dowsley, Martine De CockICML 2021 · 25 citations
- PriFU: Capturing Task-Relevant Information Without Adversarial LearningXiuli Bi, Yang Hu, Bo Liu, Weisheng Li et al.ACM MM 2024
- Secure Quantized Training for Deep LearningMarcel Keller, Ke SunICML 2022 · 84 citations
