Learning discrete distributions: user vs item-level privacy
Yuhan Liu, Ananda Theertha Suresh, Felix X. Yu, Sanjiv Kumar, Michael Riley
摘要
Much of the literature on differential privacy focuses on item-level privacy, where loosely speaking, the goal is to provide privacy per item or training example. However, recently many practical applications such as federated learning require preserving privacy for all items of a single user, which is much harder to achieve. Therefore understanding the theoretical limit of user-level privacy becomes crucial. We study the fundamental problem of learning discrete distributions over symbols with user-level differential privacy. If each user has samples, we show that straightforward applications of Laplace or Gaussian mechanisms require the number of users to be to achieve an distance of between the true and estimated distributions, with the privacy-induced penalty independent of the number of samples per user . Moreover, we show that any mechanism that only operates on the final aggregate should require a user complexity of the same order. We then propose a mechanism such that the number of users scales as and further show that it is nearly-optimal under certain regimes. Thus the privacy penalty is times smaller compared to the standard mechanisms. We also propose general techniques for obtaining lower bounds on restricted differentially private estimators and a lower bound on the total variation between binomial distributions, both of which might be of independent interest.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Learning with User-Level PrivacyDaniel Levy, Ziteng Sun, Kareem Amin, Satyen Kale 等NeurIPS 2021 · 被引用 113 次
- User-Level Differentially Private Learning via Correlated SamplingBadih Ghazi, Ravi Kumar, Pasin ManurangsiNeurIPS 2021 · 被引用 45 次
- New Lower Bounds for Private Estimation and a Generalized Fingerprinting LemmaGautam Kamath, Argyris Mouzakis, Vikrant SinghalNeurIPS 2022 · 被引用 41 次
- Mean Estimation with User-level Privacy under Data HeterogeneityRachel Cummings, Vitaly Feldman, Audra McMillan, Kunal TalwarNeurIPS 2022 · 被引用 35 次
- Tight and Robust Private Mean Estimation with Few UsersShyam Narayanan, Vahab S. Mirrokni, Hossein EsfandiariICML 2022 · 被引用 34 次
它引用的顶会 Paper1
相关 Paper
- User-Level Differential Privacy With Few Examples Per UserBadih Ghazi, Pritish Kamath, Ravi Kumar, Pasin Manurangsi 等NeurIPS 2023 · 被引用 19 次
- Improved Bounds for Pure Private Agnostic Learning: Item-Level and User-Level PrivacyBo Li, Wei Wang, Peng YeICML 2024 · 被引用 1 次
- The Poisson Binomial Mechanism for Unbiased Federated Learning with Secure AggregationWei-Ning Chen, Ayfer Özgür, Peter KairouzICML 2022 · 被引用 57 次
- Sample-Efficient Private Learning of Mixtures of GaussiansHassan Ashtiani, Mahbod Majid, Shyam NarayananNeurIPS 2024
- Approximate Differential Privacy of the ℓ2 MechanismMatthew Joseph, Alex Kulesza, Alexander YuICML 2025
