Geometry-Misalignment in Distributional Learning
Tao Wang, Xiaoting Zhong
摘要
Distributional learning problems optimize discrepancies between probability measures, including optimal transport or Sinkhorn divergence, yet are typically optimized using Euclidean first-order methods in parameter space. We show this mismatch is structural rather than algorithmic. We introduce geometry-misalignment, a local condition number that measures distortion between Euclidean geometry and the intrinsic geometry induced by a distributional objective. For a broad class of problems, we establish lower bounds demonstrating that Euclidean first-order methods incur an unavoidable convergence slowdown proportional to misalignment, even under intrinsic strong convexity and smoothness. We further prove geometry-aware preconditioned methods attain matching upper bounds independent of misalignment, yielding a sharp separation between Euclidean optimization and geometry-aware optimization. Beyond convergence rates, we show geometry-misalignment induces an optimization-dependent excess risk term under finite budgets, directly linking optimization geometry with statistical efficiency. We develop a geometry-calibrated optimization framework that estimates misalignment and selectively activates geometry-aware updates when necessary. Experiments on distribution matching for domain adaptation validate the theory, with improvements concentrated in high-misalignment regimes and negligible overhead.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Stability of Stochastic Gradient Descent on Nonsmooth Convex LossesRaef Bassily, Vitaly Feldman, Cristóbal Guzmán, Kunal TalwarNeurIPS 2020 · 被引用 240 次
- A General Approach to Fairness with Optimal TransportSilvia Chiappa, Ray Jiang, Tom Stepleton, Aldo Pacchiano 等AAAI 2020 · 被引用 94 次
- Kronecker-Factored Approximate Curvature for Physics-Informed Neural NetworksFelix Dangel, Johannes Müller, Marius ZeinhoferNeurIPS 2024 · 被引用 31 次
- Vector Quantized Wasserstein Auto-EncoderLong Tung Vuong, Trung Le, He Zhao, Chuanxia Zheng 等ICML 2023 · 被引用 24 次
- Efficient preconditioned stochastic gradient descent for estimation in latent variable modelsCharlotte Baey, Maud Delattre, Estelle Kuhn, Jean-Benoist Leger 等ICML 2023 · 被引用 6 次
相关 Paper
- GeoDM: Geometry-aware Distribution Matching for Dataset DistillationXuhui Li, Zhengquan Luo, Zihui Cui, Kai Zhao 等ICML 2026 · 被引用 2 次
- Adversarial Support AlignmentShangyuan Tong, Timur Garipov, Yang Zhang, Shiyu Chang 等ICLR 2022 · 被引用 10 次
- Convex Distance Operator Transport: A Convex and Geometry-Preserving FormulationJunhyoung Chung, Euijong Song, Won Hwa Kim, Gunwoong ParkICML 2026
- Incorporating Importance Weighting in Optimal Transport Based Domain AlignmentOkan Koç, Alexander Soen, Shanglin Li, Masashi SugiyamaICML 2026
- Geometric Dataset Distances via Optimal TransportDavid Alvarez-Melis, Nicolò FusiNeurIPS 2020 · 被引用 267 次
