Practical Lossless Federated Singular Vector Decomposition over Billion-Scale Data
Di Chai, Leye Wang, Junxue Zhang, Liu Yang, Shuowei Cai, Kai Chen, Qiang Yang
Abstract
With the enactment of privacy-preserving regulations, e.g., GDPR, federated SVD is proposed to enable SVD-based applications over different data sources without revealing the original data. However, many SVD-based applications cannot be well supported by existing federated SVD solutions. The crux is that these solutions, adopting either differential privacy (DP) or homomorphic encryption (HE), suffer from accuracy loss caused by unremovable noise or degraded efficiency due to inflated data. In this paper, we propose FedSVD, a practical lossless federated SVD method over billion-scale data, which can simultaneously achieve lossless accuracy and high efficiency. At the heart of FedSVD is a lossless matrix masking scheme delicately designed for SVD: 1) While adopting the masks to protect private data, FedSVD completely removes them from the final results of SVD to achieve lossless accuracy; and 2) As the masks do not inflate the data, FedSVD avoids extra computation and communication overhead during the factorization to maintain high efficiency. Experiments with real-world datasets show that FedSVD is over 10000× faster than the HE-based method and has 10 orders of magnitude smaller error than the DP-based solution (𝜖 = 0.1, 𝛿 = 0.1) on SVD tasks. We further build and evaluate FedSVD over three real-world applications: principal components analysis (PCA), linear regression (LR), and latent semantic analysis (LSA), to show its superior performance in practice. On federated LR tasks, compared with two state-of-the-art solutions: FATE [17] and SecureML [19], FedSVD-LR is 100× faster than SecureML and 10× faster than FATE. CCS Concepts • Security and privacy → Privacy-preserving protocols; • Computing methodologies → Factorization methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c01b6424-d9e3-45a1-a2da-06b6829b6c48Cited by top-tier papers7
- COLA: Cross-city Mobility Transformer for Human Trajectory SimulationYu Wang, Tongya Zheng, Yuxuan Liang, Shunyu Liu et al.WWW 2024 · 37 citations
- FLASH: Towards a High-performance Hardware Acceleration Architecture for Cross-silo Federated LearningJunxue Zhang, Xiaodian Cheng, Wei Wang, Liu Yang et al.NSDI 2023 · 29 citations
- Efficient Decentralized Federated Singular Vector DecompositionDi Chai, Junxue Zhang, Liu Yang, Yilun Jin et al.USENIX ATC 2024 · 9 citations
- Accelerating Secure Collaborative Machine Learning with Protocol-Aware RDMAZhenghang Ren, Mingxuan Fan, Zilong Wang, Junxue Zhang et al.USENIX Security 2024 · 6 citations
- Vertical Federated Learning with Missing Features During Training and InferencePedro Valdeira, Shiqiang Wang, Yuejie ChiICLR 2025
Builds on4
- Practical Secure Aggregation for Privacy-Preserving Machine LearningKallista A. Bonawitz, Vladimir Ivanov, Ben Kreuter, Antonio Marcedone et al.CCS 2017 · 3,936 citations
- SecureML: A System for Scalable Privacy-Preserving Machine LearningPayman Mohassel, Yupeng ZhangS&P 2017 · 2,107 citations
- Federated Principal Component AnalysisAndreas Grammenos, Rodrigo Mendoza-Smith, Jon Crowcroft, Cecilia MascoloNeurIPS 2020 · 85 citations
- Sphinx: Enabling Privacy-Preserving Online Learning over the CloudHan Tian, Chaoliang Zeng, Zhenghang Ren, Di Chai et al.S&P 2022 · 37 citations
Related papers
- FedSVD: Adaptive Orthogonalization for Private Federated Learning with LoRASeanie Lee, Sangwoo Park, Dong Bok Lee, Dominik Wagner et al.NeurIPS 2025 · 18 citations
- Scalable and Privacy-Preserving Federated Principal Component AnalysisDavid Froelicher, Hyunghoon Cho, Manaswitha Edupalli, Joao Sa Sousa et al.S&P 2023
- Efficient Heterogeneity-Aware Federated Active Data SelectionYing-Peng Tang, Chao Ren, Xiaoli Tang, Sheng-Jun Huang et al.ICML 2025
- Banded Square Root Matrix Factorization for Differentially Private Model TrainingNikita P. Kalinin, Christoph H. LampertNeurIPS 2024 · 18 citations
- Differentially Private Federated Low Rank Adaptation Beyond Fixed-MatrixMing Wen, Jiaqi Zhu, Yuedong Xu, Yipeng Zhou et al.NeurIPS 2025 · 6 citations
