Federated Latent Dirichlet Allocation: A Local Differential Privacy Based Framework
Yansheng Wang, Yongxin Tong, Dingyuan Shi
摘要
Latent Dirichlet Allocation (LDA) is a widely adopted topic model for industrial-grade text mining applications. However, its performance heavily relies on the collection of large amount of text data from users' everyday life for model training. Such data collection risks severe privacy leakage if the data collector is untrustworthy. To protect text data privacy while allowing accurate model training, we investigate federated learning of LDA models. That is, the model is collaboratively trained between an untrustworthy data collector and multiple users, where raw text data of each user are stored locally and not uploaded to the data collector. To this end, we propose FedLDA, a local differential privacy (LDP) based framework for federated learning of LDA models. Central in FedLDA is a novel LDP mechanism called Random Response with Priori (RRP), which provides theoretical guarantees on both data privacy and model accuracy. We also design techniques to reduce the communication cost between the data collector and the users during model training. Extensive experiments on three open datasets verified the effectiveness of our solution.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Provably Secure Federated Learning against Malicious ClientsXiaoyu Cao, Jinyuan Jia, Neil Zhenqiang GongAAAI 2021 · 被引用 161 次
- Hierarchical Personalized Federated Learning for User ModelingJinze Wu, Qi Liu, Zhenya Huang, Yuting Ning 等WWW 2021 · 被引用 97 次
- An Efficient Approach for Cross-Silo Federated Learning to RankYansheng Wang, Yongxin Tong, Dingyuan Shi, Ke XuICDE 2021 · 被引用 36 次
- Distribution-Regularized Federated Learning on Non-IID DataYansheng Wang, Yongxin Tong, Zimu Zhou, Ruisheng Zhang 等ICDE 2023 · 被引用 31 次
- Prompt-enhanced Federated Content Representation Learning for Cross-domain RecommendationLei Guo, Ziang Lu, Junliang Yu, Quoc Viet Hung Nguyen 等WWW 2024 · 被引用 30 次
它引用的顶会 Paper3
- Practical Secure Aggregation for Privacy-Preserving Machine LearningKallista A. Bonawitz, Vladimir Ivanov, Ben Kreuter, Antonio Marcedone 等CCS 2017 · 被引用 3,936 次
- Heavy Hitter Estimation over Set-Valued Data with Local Differential PrivacyZhan Qin, Yin Yang, Ting Yu, Issa Khalil 等CCS 2016 · 被引用 344 次
- Locally Differentially Private Frequent Itemset MiningTianhao Wang, Ninghui Li, Somesh JhaS&P 2018 · 被引用 196 次
相关 Paper
- LabelDP-Pro: Learning with Label Differential Privacy via ProjectionsBadih Ghazi, Yangsibo Huang, Pritish Kamath, Ravi Kumar 等ICLR 2024 · 被引用 4 次
- Deep Learning with Label Differential PrivacyBadih Ghazi, Noah Golowich, Ravi Kumar, Pasin Manurangsi 等NeurIPS 2021 · 被引用 193 次
- Enhancing Privacy Preservation in Federated Learning via Learning Rate PerturbationGuangnian Wan, Haitao Du, Xuejing Yuan, Jun Yang 等ICCV 2023 · 被引用 2 次
- Efficient and Differentially Private Federated LLM Fine-Tuning on Heterogeneous ClientsNan Yan, Yuqing Li, Xiong Wang, Jing Chen 等KDD 2026
- Differentially Private Federated Low Rank Adaptation Beyond Fixed-MatrixMing Wen, Jiaqi Zhu, Yuedong Xu, Yipeng Zhou 等NeurIPS 2025 · 被引用 6 次
