Rethinking Losses for Diffusion Bridge Samplers
Sebastian Sanokowski, Lukas Gruber, Christoph Bartmann, Sepp Hochreiter, Sebastian Lehner
摘要
Diffusion bridges are a promising class of deep-learning methods for sampling from unnormalized distributions. Recent works show that the Log Variance (LV) loss consistently outperforms the reverse Kullback-Leibler (rKL) loss when using the reparametrization trick to compute rKL-gradients. While the on-policy LV loss yields identical gradients to the rKL loss when combined with the log-derivative trick for diffusion samplers with non-learnable forward processes, this equivalence does not hold for diffusion bridges or when diffusion coefficients are learned. Based on this insight we argue that for diffusion bridges the LV loss does not represent an optimization objective that can be motivated like the rKL loss via the data processing inequality. Our analysis shows that employing the rKL loss with the log-derivative trick (rKL-LD) does not only avoid these conceptual problems but also consistently outperforms the LV loss. Experimental results with different types of diffusion bridges on challenging benchmarks show that samplers trained with the rKL-LD loss achieve better performance. From a practical perspective we find that rKL-LD requires significantly less hyperparameter optimization and yields more stable training behavior. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Trust Region Constrained Measure Transport in Path Space for Stochastic Optimal Control and InferenceDenis Blessing, Julius Berner, Lorenz Richter, Carles Domingo-Enrich 等NeurIPS 2025 · 被引用 24 次
- Bridge Matching Sampler: Scalable Sampling via Generalized Fixed-Point Diffusion MatchingDenis Blessing, Lorenz Richter, Julius Berner, Egor Malitskiy 等ICML 2026 · 被引用 8 次
- Discrete Adjoint Schrödinger Bridge SamplerWei Guo, Yuchen Zhu, Xiaochen Du, Juno Nam 等ICML 2026 · 被引用 3 次
- Discrete Diffusion Samplers and Bridges: Off-Policy Algorithms and Applications in Latent SpacesArran Carter, Sanghyeok Choi, Kirill Tamogashev, Víctor Elvira 等ICML 2026 · 被引用 1 次
- Minimax-Optimal Aggregation for Density Ratio EstimationLukas Gruber, Markus Holzleitner, Sepp Hochreiter, Werner ZellingerICLR 2026
它引用的顶会 Paper21
- On the Variance of the Adaptive Learning Rate and BeyondLiyuan Liu, Haoming Jiang, Pengcheng He, Weizhu Chen 等ICLR 2020 · 被引用 2,210 次
- Maximum Likelihood Training of Score-Based Diffusion ModelsYang Song, Conor Durkan, Iain Murray, Stefano ErmonNeurIPS 2021 · 被引用 958 次
- Diffusion Schrödinger Bridge with Applications to Score-Based Generative ModelingValentin De Bortoli, James Thornton, Jeremy Heng, Arnaud DoucetNeurIPS 2021 · 被引用 811 次
- Torsional Diffusion for Molecular Conformer GenerationBowen Jing, Gabriele Corso, Jeffrey Chang, Regina Barzilay 等NeurIPS 2022 · 被引用 413 次
- Path Integral Sampler: A Stochastic Control Approach For SamplingQinsheng Zhang, Yongxin ChenICLR 2022 · 被引用 177 次
相关 Paper
- Improved sampling via learned diffusionsLorenz Richter, Julius BernerICLR 2024 · 被引用 103 次
- A Diffusion Model Framework for Unsupervised Neural Combinatorial OptimizationSebastian Sanokowski, Sepp Hochreiter, Sebastian LehnerICML 2024 · 被引用 60 次
- Underdamped Diffusion Bridges with Applications to SamplingDenis Blessing, Julius Berner, Lorenz Richter, Gerhard NeumannICLR 2025
- Neural Guided Diffusion BridgesGefan Yang, Frank van der Meulen, Stefan SommerICML 2025
- Neural Flow Diffusion Models: Learnable Forward Process for Improved Diffusion ModellingGrigory Bartosh, Dmitry P. Vetrov, Christian Andersson NaessethNeurIPS 2024 · 被引用 49 次
