Hash-Prune-Invert: Improved Differentially Private Heavy-Hitter Detection in the Two-Server Model
Borja Balle, James Bell-Clark, Albert Cheu, Adrià Gascón, Jonathan Katz, Mariana Raykova, Phillipp Schoppmann, Thomas Steinke
Abstract
Differentially private (DP) heavy-hitter detection is an important primitive for data analysis. Given a threshold <tex></tex> and a dataset of <tex></tex> items from a domain of size <tex></tex>, such detection algorithms ignore items occurring fewer than <tex></tex> times while identifying items occurring more than <tex></tex> times; we call <tex></tex> the error margin. In the central model where a curator holds the entire dataset, <tex></tex>-DP algorithms can achieve error margin <tex></tex>, which is optimal when <tex></tex>. Several works, e.g., Poplar (S&P 2021), have proposed protocols in which two or more non-colluding servers jointly compute the heavy hitters from inputs held by <tex></tex> clients. Unfortunately, existing protocols suffer from an undesirable dependence on Iog <tex></tex> in terms of both server efficiency (computation, communication, and round complexity) and accuracy (i.e., error margin), making them unsuitable for large domains (e.g., when items are kB-long strings, log <tex></tex>). We present hash-prune-invert (HPI), a technique for compiling any heavy-hitter protocol with the log <tex></tex> dependencies mentioned above into a new protocol with improvements across the board: computation, communication, and round complexity depend (roughly) on log <tex></tex> rather than log <tex></tex>, and the error margin is independent of <tex></tex>. Our transformation preserves privacy against an active adversary corrupting at most one of the servers and any number of clients. We apply HPI to an improved version of Poplar, also introduced in this work, that improves Poplar's error margin by roughly a factor of <tex></tex> (regardless of <tex></tex>. Our experiments confirm that the resulting protocol improves efficiency and accuracy for large <tex></tex>.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4d9e36ae-bcef-4aec-8a6b-9fcd28c158a1Cited by top-tier papers2
- Piquant: Private Quantile Estimation in the Two-Server ModelHannah Keller, Jacob Imola, Fabrizio Boninsegna, Rasmus Pagh et al.CCS 2026
- Efficient, Secure, Differentially Private Deep Learning in the Two-Server ModelJun Feng, Hong Sun, Pengfei Zhang, Bocheng Ren et al.AAAI 2026
Builds on5
- Function Secret Sharing: Improvements and ExtensionsElette Boyle, Niv Gilboa, Yuval IshaiCCS 2016 · 404 citations
- Authenticated Garbling and Efficient Maliciously Secure Two-Party ComputationXiao Wang, Samuel Ranellucci, Jonathan KatzCCS 2017 · 212 citations
- Lightweight Techniques for Private Heavy HittersDan Boneh, Elette Boyle, Henry Corrigan-Gibbs, Niv Gilboa et al.S&P 2021 · 134 citations
- Distributed, Private, Sparse Histograms in the Two-Server ModelJames Bell, Adrià Gascón, Badih Ghazi, Ravi Kumar et al.CCS 2022 · 19 citations
- Precio: Private Aggregate Measurement via Oblivious ShufflingErik Anderson, Melissa Chase, F. Betül Durak, Kim Laine et al.CCS 2024 · 3 citations
Related papers
- POPSTAR: Lightweight Threshold Reporting with Reduced LeakageHanjun Li, Sela Navot, Stefano TessaroUSENIX Security 2024 · 5 citations
- Frequency Estimation in the Shuffle Model with Almost a Single MessageQiyao Luo, Yilei Wang, Ke YiCCS 2022 · 6 citations
- Secure Multi-party Computation of Differentially Private Heavy HittersJonas Böhler, Florian KerschbaumCCS 2021 · 34 citations
- Heavy Hitter Estimation over Set-Valued Data with Local Differential PrivacyZhan Qin, Yin Yang, Ting Yu, Issa Khalil et al.CCS 2016 · 344 citations
- An Iconic Heavy Hitters Algorithm Made PrivateRayne HollandCCS 2026
