A Huber Loss Minimization Approach to Mean Estimation under User-level Differential Privacy
Puning Zhao, Lifeng Lai, Li Shen, Qingming Li, Jiafei Wu, Zhe Liu
摘要
Privacy protection of users' entire contribution of samples is important in distributed systems. The most effective approach is the two-stage scheme, which finds a small interval first and then gets a refined estimate by clipping samples into the interval. However, the clipping operation induces bias, which is serious if the sample distribution is heavy-tailed. Besides, users with large local sample sizes can make the sensitivity much larger, thus the method is not suitable for imbalanced users. Motivated by these challenges, we propose a Huber loss minimization approach to mean estimation under user-level differential privacy. The connecting points of Huber loss can be adaptively adjusted to deal with imbalanced users. Moreover, it avoids the clipping operation, thus significantly reducing the bias compared with the two-stage approach. We provide a theoretical analysis of our approach, which gives the noise strength needed for privacy protection, as well as the bound of mean squared error. The result shows that the new method is much less sensitive to the imbalance of user-wise sample sizes and the tail of sample distributions. Finally, we perform numerical experiments to validate our theoretical analysis.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Private Mean Estimation with Person-Level Differential PrivacySushant Agarwal, Gautam Kamath, Mahbod Majid, Argyris Mouzakis 等SODA 2025 · 被引用 6 次
- Locally Optimal Private Sampling: Beyond the Global MinimaxHrad Ghoukasian, Bonwoo Lee, Shahab AsoodehNeurIPS 2025 · 被引用 2 次
- Consistent Estimation of Numerical Distributions Under Local Differential Privacy by Wavelet ExpansionPuning Zhao, Zhikun Zhang, Bo Sun, Li Shen 等S&P 2026 · 被引用 2 次
- Differential Private Stochastic Optimization with Heavy-tailed Data: Towards Optimal RatesPuning Zhao, Jiafei Wu, Zhe Liu, Chong Wang 等AAAI 2025 · 被引用 1 次
- Hephaestus: Mixture Generative Modeling with Energy Guidance for Large-scale QoS DegradationNguyen Do, Bach Ngo, Youval Kashuv, Canh V. Pham 等NeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper27
- Deep Learning with Differential PrivacyMartín Abadi, Andy Chu, Ian J. Goodfellow, H. Brendan McMahan 等CCS 2016 · 被引用 7,620 次
- Fair Resource Allocation in Federated LearningTian Li, Maziar Sanjabi, Ahmad Beirami, Virginia SmithICLR 2020 · 被引用 971 次
- Why are Adaptive Methods Good for Attention Models?Jingzhao Zhang, Sai Praneeth Karimireddy, Andreas Veit, Seungyeon Kim 等NeurIPS 2020 · 被引用 397 次
- The Heavy-Tail Phenomenon in SGDMert Gürbüzbalaban, Umut Simsekli, Lingjiong ZhuICML 2021 · 被引用 165 次
- Learning with User-Level PrivacyDaniel Levy, Ziteng Sun, Kareem Amin, Satyen Kale 等NeurIPS 2021 · 被引用 113 次
相关 Paper
- Algorithms for bounding contribution for histogram estimation under user-level privacyYuhan Liu, Ananda Theertha Suresh, Wennan Zhu, Peter Kairouz 等ICML 2023 · 被引用 14 次
- Privacy for Free: Leveraging Local Differential Privacy Perturbed Data from Multiple ServicesRong Du, Qingqing Ye, Yue Fu, Haibo HuVLDB 2025 · 被引用 2 次
- Mean Estimation with User-level Privacy under Data HeterogeneityRachel Cummings, Vitaly Feldman, Audra McMillan, Kunal TalwarNeurIPS 2022 · 被引用 35 次
- Instance-optimal Mean Estimation Under Differential PrivacyZiyue Huang, Yuting Liang, Ke YiNeurIPS 2021 · 被引用 74 次
- On Private and Robust BanditsYulian Wu, Xingyu Zhou, Youming Tao, Di WangNeurIPS 2023 · 被引用 12 次
