Differentially Private Adaptive Optimization with Delayed Preconditioners
Tian Li, Manzil Zaheer, Ken Liu, Sashank J. Reddi, Hugh Brendan McMahan, Virginia Smith
Abstract
Privacy noise may negate the benefits of using adaptive optimizers in differentially private model training. Prior works typically address this issue by using auxiliary information (e.g., public data) to boost the effectiveness of adaptive optimization. In this work, we explore techniques to estimate and efficiently adapt to gradient geometry in private adaptive optimization without auxiliary data. Motivated by the observation that adaptive methods can tolerate stale preconditioners, we propose differentially private adaptive training with delayed preconditioners (DP^2), a simple method that constructs delayed but less noisy preconditioners to better realize the benefits of adaptivity. Theoretically, we provide convergence guarantees for our method for both convex and non-convex problems, and analyze trade-offs between delay and privacy noise reduction. Empirically, we explore DP^2 across several real-world datasets, demonstrating that it can improve convergence speed by as much as 4x relative to non-adaptive baselines and match the performance of state-of-the-art optimization methods that require auxiliary data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8df20366-e272-417a-bdfc-5948eaf3591dCited by top-tier papers12
- Dynamic Personalized Federated Learning with Adaptive Differential PrivacyXiyuan Yang, Wenke Huang, Mang YeNeurIPS 2023 · 166 citations
- DP-AdamBC: Your DP-Adam Is Actually DP-SGD (Unless You Apply Bias Correction)Qiaoyue Tang, Frederick Shpilevskiy, Mathias LécuyerAAAI 2024 · 33 citations
- Practical Differentially Private Hyperparameter Tuning with SubsamplingAntti Koskela, Tejas D. KulkarniNeurIPS 2023 · 32 citations
- A New Linear Scaling Rule for Private Adaptive Hyperparameter OptimizationAshwinee Panda, Xinyu Tang, Saeed Mahloujifar, Vikash Sehwag et al.ICML 2024 · 15 citations
- Efficient Adaptive Federated OptimizationSu Hyeong Lee, Sidharth Sharma, Manzil Zaheer, Tian LiNeurIPS 2025 · 6 citations
Builds on11
- Deep Learning with Differential PrivacyMartín Abadi, Andy Chu, Ian J. Goodfellow, H. Brendan McMahan et al.CCS 2016 · 7,620 citations
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett et al.ICLR 2021 · 1,917 citations
- Why are Adaptive Methods Good for Attention Models?Jingzhao Zhang, Sai Praneeth Karimireddy, Andreas Veit, Seungyeon Kim et al.NeurIPS 2020 · 397 citations
- Practical and Private (Deep) Learning Without Sampling or ShufflingPeter Kairouz, Brendan McMahan, Shuang Song, Om Thakkar et al.ICML 2021 · 239 citations
- Bypassing the Ambient Dimension: Private SGD with Gradient Subspace IdentificationYingxue Zhou, Steven Wu, Arindam BanerjeeICLR 2021 · 118 citations
Related papers
- Private Adaptive Optimization with Side informationTian Li, Manzil Zaheer, Sashank J. Reddi, Virginia SmithICML 2022 · 46 citations
- DP-KFC: Data-Free Preconditioning for Privacy-Preserving Deep LearningMarc Molina Van den bosch, Riccardo Taiello, Albert Aillet, Andrea Protani et al.ICML 2026
- DOPPLER: Differentially Private Optimizers with Low-pass Filter for Privacy Noise ReductionXinwei Zhang, Zhiqi Bu, Mingyi Hong, Meisam RazaviyaynNeurIPS 2024 · 10 citations
- GeoClip: Geometry-Aware Clipping for Differentially Private SGDAtefeh Gilani, Naima Tasnim, Lalitha Sankar, Oliver KosutNeurIPS 2025 · 5 citations
- FIBER: A Differentially Private Optimizer with Filter-Aware Innovation Bias CorrectionMINH DUC DO, Thao Do, Minh Hoang, Anh Le Duc Tran et al.ICML 2026 · 1 citation
