Differentiable Sorting Networks for Scalable Sorting and Ranking Supervision
Felix Petersen, Christian Borgelt, Hilde Kuehne, Oliver Deussen
摘要
Sorting and ranking supervision is a method for training neural networks end-to-end based on ordering constraints. That is, the ground truth order of sets of samples is known, while their absolute values remain unsupervised. For that, we propose differentiable sorting networks by relaxing their pairwise conditional swap operations. To address the problems of vanishing gradients and extensive blurring that arise with larger numbers of layers, we propose mapping activations to regions with moderate gradients. We consider odd-even as well as bitonic sorting networks, which outperform existing relaxations of the sorting operation. We show that bitonic sorting networks can achieve stable training on large input sets of up to 1024 elements.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- Deep Differentiable Logic Gate NetworksFelix Petersen, Christian Borgelt, Hilde Kuehne, Oliver DeussenNeurIPS 2022 · 被引用 117 次
- Differentiable Top-k Classification LearningFelix Petersen, Hilde Kuehne, Christian Borgelt, Oliver DeussenICML 2022 · 被引用 48 次
- Fast, Differentiable and Sparse Top-k: a Convex Analysis PerspectiveMichael Eli Sander, Joan Puigcerver, Josip Djolonga, Gabriel Peyré 等ICML 2023 · 被引用 35 次
- Git Re-Basin: Merging Models modulo Permutation SymmetriesSamuel K. Ainsworth, Jonathan Hayase, Siddhartha S. SrinivasaICLR 2023 · 被引用 32 次
- Monotonic Differentiable Sorting NetworksFelix Petersen, Christian Borgelt, Hilde Kuehne, Oliver DeussenICLR 2022 · 被引用 32 次
它引用的顶会 Paper2
相关 Paper
- Generalized Neural Sorting Networks with Error-Free Differentiable Swap FunctionsJungtaek Kim, Jeongbeen Yoon, Minsu ChoICLR 2024 · 被引用 5 次
- OPS: An Order-Preserving Sorting Network for Information RetrievalChao Wang, Yongxiang Tang, Guikai Luan, Kaiyuan Li 等SIGIR 2026
- Newton Losses: Using Curvature Information for Learning with Differentiable AlgorithmsFelix Petersen, Christian Borgelt, Tobias Sutter, Hilde Kuehne 等NeurIPS 2024 · 被引用 3 次
- Learning with Algorithmic Supervision via Continuous RelaxationsFelix Petersen, Christian Borgelt, Hilde Kuehne, Oliver DeussenNeurIPS 2021 · 被引用 33 次
- Differentiable sorting for censored time-to-event dataAndre Vauvelle, Benjamin Wild, Roland Eils, Spiros C. DenaxasNeurIPS 2023 · 被引用 6 次
