Lune

ICLR2026顶会

Adaptive Methods Are Preferable in High Privacy Settings: An SDE Perspective

Enea Monzio Compagnoni, Alessandro Stanghellini, Rustem Islamov, Aurélien Lucchi, Anastasia Koloskova

2026年份
2被引次数

摘要

Differential Privacy (DP) is becoming central to large-scale training as privacy regulations tighten. We revisit how DP noise interacts with adaptivity in optimization through the lens of stochastic differential equations, providing the first SDE-based analysis of private optimizers. Focusing on DP-SGD and DP-SignSGD under per-example clipping, we show a sharp contrast under fixed hyperparameters: DP-SGD converges at a Privacy-Utility Trade-Off of O(1/ε2)\mathcal{O}(1/\varepsilon^2) with speed independent of ε\varepsilon, while DP-SignSGD converges at a speed linear in ε\varepsilon with a O(1/ε)\mathcal{O}(1/\varepsilon) trade-off, dominating in high-privacy or large batch noise regimes. By contrast, under optimal learning rates, both methods achieve comparable theoretical asymptotic performance; however, the optimal learning rate of DP-SGD scales linearly with ε\varepsilon, while that of DP-SignSGD is essentially ε\varepsilon-independent. This makes adaptive methods far more practical, as their hyperparameters transfer across privacy levels with little or no re-tuning. Empirical results confirm our theory across training and test metrics, and empirically extend from DP-SignSGD to DP-Adam.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

lune papers fulltext cb1990ce-ba33-440a-92be-e46001490bb4

它引用的顶会 Paper22

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖