Flowing Datasets with Wasserstein over Wasserstein Gradient Flows
Clément Bonet, Christophe Vauthier, Anna Korba
摘要
Many applications in machine learning involve data represented as probability distributions. The emergence of such data requires radically novel techniques to design tractable gradient flows on probability distributions over this type of (infinitedimensional) objects. For instance, being able to flow labeled datasets is a core task for applications ranging from domain adaptation to transfer learning or dataset distillation. In this setting, we propose to represent each class by the associated conditional distribution of features, and to model the dataset as a mixture distribution supported on these classes (which are themselves probability distributions), meaning that labeled datasets can be seen as probability distributions over probability distributions. We endow this space with a metric structure from optimal transport, namely the Wasserstein over Wasserstein (WoW) distance, derive a differential structure on this space, and define WoW gradient flows. The latter enables to design dynamics over this space that decrease a given objective functional. We apply our framework to transfer learning and dataset distillation tasks, leveraging our gradient flow construction as well as novel tractable functionals that take the form of Maximum Mean Discrepancies with Sliced-Wasserstein based kernels between probability distributions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Learning of Population Dynamics: Inverse Optimization Meets JKO SchemeMikhail Persiianov, Jiawei Chen, Petr Mokrov, Alexander Tyurin 等ICLR 2026 · 被引用 7 次
- Slicing Wasserstein over Wasserstein via Functional Optimal TransportMoritz Piening, Robert BeinertICLR 2026 · 被引用 5 次
- A Novel Sliced Fused Gromov-Wasserstein DistanceMoritz Piening, Robert BeinertAAAI 2026 · 被引用 3 次
- Splat Regression ModelsMara Daniels, Philippe RigolletICLR 2026 · 被引用 2 次
- Dense associative memory for Gaussian distributionsChandan Tankala, Krishna BalasubramanianICML 2026 · 被引用 1 次
它引用的顶会 Paper19
- Dataset Condensation with Differentiable Siamese AugmentationBo Zhao, Hakan BilenICML 2021 · 被引用 390 次
- Geometric Dataset Distances via Optimal TransportDavid Alvarez-Melis, Nicolò FusiNeurIPS 2020 · 被引用 267 次
- Variational inference via Wasserstein gradient flowsMarc Lambert, Sinho Chewi, Francis R. Bach, Silvère Bonnabel 等NeurIPS 2022 · 被引用 123 次
- The Wasserstein Proximal Gradient AlgorithmAdil Salim, Anna Korba, Giulia LuiseNeurIPS 2020 · 被引用 74 次
- Averaging on the Bures-Wasserstein manifold: dimension-free convergence of gradient descentJason M. Altschuler, Sinho Chewi, Patrik Gerber, Austin J. StrommeNeurIPS 2021 · 被引用 60 次
相关 Paper
- Dataset Dynamics via Gradient Flows in Probability SpaceDavid Alvarez-Melis, Nicolò FusiICML 2021 · 被引用 24 次
- Distribution Regression with Sliced Wasserstein KernelsDimitri Meunier, Massimiliano Pontil, Carlo CilibertoICML 2022 · 被引用 24 次
- Differentiable Generalized Sliced Wasserstein PlansLaetitia Chapel, Romain Tavenard, Samuel VaiterNeurIPS 2025 · 被引用 11 次
- Diffeomorphic Mesh Deformation via Efficient Optimal Transport for Cortical Surface ReconstructionThanh-Tung Le, Khai Nguyen, Shanlin Sun, Kun Han 等ICLR 2024 · 被引用 9 次
- Wasserstein Transfer LearningKaicheng Zhang, Sinian Zhang, Doudou Zhou, Yidong ZhouNeurIPS 2025 · 被引用 2 次
