Backpropagation through Combinatorial Algorithms: Identity with Projection Works
Subham Sekhar Sahoo, Anselm Paulus, Marin Vlastelica, Vít Musil, Volodymyr Kuleshov, Georg Martius
摘要
Embedding discrete solvers as differentiable layers has given modern deep learning architectures combinatorial expressivity and discrete reasoning capabilities. The derivative of these solvers is zero or undefined, therefore a meaningful replacement is crucial for effective gradient-based learning. Prior works rely on smoothing the solver with input perturbations, relaxing the solver to continuous problems, or interpolating the loss landscape with techniques that typically require additional solver calls, introduce extra hyper-parameters, or compromise performance. We propose a principled approach to exploit the geometry of the discrete solution space to treat the solver as a negative identity on the backward pass and further provide a theoretical justification. Our experiments demonstrate that such a straightforward hyperparameter-free approach is able to compete with previous more complex methods on numerous experiments such as backpropagation through discrete samplers, deep graph matching, and image retrieval. Furthermore, we substitute the previously proposed problem-specific and label-dependent margin with a generic regularization procedure that prevents cost collapse and increases robustness. Code is available at github.com/martius-lab/solver-differentiation-identity.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Simple and Effective Masked Diffusion Language ModelsSubham S. Sahoo, Marianne Arriola, Yair Schiff, Aaron Gokaslan 等NeurIPS 2024 · 被引用 929 次
- Diffusion Models With Learned Adaptive NoiseSubham S. Sahoo, Aaron Gokaslan, Christopher De Sa, Volodymyr KuleshovNeurIPS 2024 · 被引用 64 次
- A-NeSI: A Scalable Approximate Method for Probabilistic Neurosymbolic InferenceEmile van Krieken, Thiviyan Thanapalasingam, Jakub M. Tomczak, Frank van Harmelen 等NeurIPS 2023 · 被引用 62 次
- One-step differentiation of iterative algorithmsJérôme Bolte, Edouard Pauwels, Samuel VaiterNeurIPS 2023 · 被引用 36 次
- Differentiable Simulation of Hard Contacts with Soft Gradients for Learning and ControlAnselm Paulus, Andreas René Geist, Pierre Schumacher, Vít Musil 等ICLR 2026 · 被引用 7 次
它引用的顶会 Paper14
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 被引用 736 次
- Differentiation of Blackbox Combinatorial SolversMarin Vlastelica Pogancic, Anselm Paulus, Vít Musil, Georg Martius 等ICLR 2020 · 被引用 341 次
- Fast Differentiable Sorting and RankingMathieu Blondel, Olivier Teboul, Quentin Berthet, Josip DjolongaICML 2020 · 被引用 285 次
- Neural Execution of Graph AlgorithmsPetar Velickovic, Rex Ying, Matilde Padovano, Raia Hadsell 等ICLR 2020 · 被引用 192 次
- Learning with Differentiable Pertubed OptimizersQuentin Berthet, Mathieu Blondel, Olivier Teboul, Marco Cuturi 等NeurIPS 2020 · 被引用 181 次
相关 Paper
- Discrete Cycle-Consistency Based Unsupervised Deep Graph MatchingSiddharth Tourani, Muhammad Haris Khan, Carsten Rother, Bogdan SavchynskyyAAAI 2024 · 被引用 5 次
- Learning Symmetric Rules with SATNetSangho Lim, Eun-Gyeol Oh, Hongseok YangNeurIPS 2022 · 被引用 5 次
- LPGD: A General Framework for Backpropagation through Embedded Optimization LayersAnselm Paulus, Georg Martius, Vít MusilICML 2024 · 被引用 5 次
- GAMnet: Robust Feature Matching via Graph Adversarial-Matching NetworkBo Jiang, Pengfei Sun, Ziyan Zhang, Jin Tang 等ACM MM 2021 · 被引用 8 次
- CombOptNet: Fit the Right NP-Hard Problem by Learning Integer Programming ConstraintsAnselm Paulus, Michal Rolínek, Vít Musil, Brandon Amos 等ICML 2021 · 被引用 73 次
