Generalized Neural Sorting Networks with Error-Free Differentiable Swap Functions
Jungtaek Kim, Jeongbeen Yoon, Minsu Cho
摘要
Sorting is a fundamental operation of all computer systems, having been a long-standing significant research topic. Beyond the problem formulation of traditional sorting algorithms, we consider sorting problems for more abstract yet expressive inputs, e.g., multi-digit images and image fragments, through a neural sorting network. To learn a mapping from a high-dimensional input to an ordinal variable, the differentiability of sorting networks needs to be guaranteed. In this paper we define a softening error by a differentiable swap function, and develop an error-free swap function that holds a non-decreasing condition and differentiability. Furthermore, a permutation-equivariant Transformer network with multi-head attention is adopted to capture dependency between given inputs and also leverage its model capacity with self-attention. Experiments on diverse sorting benchmarks show that our methods perform better than or comparable to baseline methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Learning Distributions over Permutations and Rankings with Factorized RepresentationsDaniel Severo, Brian Karrer, Niklas NolteICLR 2026 · 被引用 1 次
- Learning Permutation Distributions via Reflected Diffusion on RanksSizhuang He, Yangtian Zhang, Shiyang Zhang, David van DijkICML 2026
- Learning to Rank by Directly Optimizing Full-Order ProbabilitiesYongxiang Tang, Chao Wang, Jincheng Lu, Yanhua Cheng 等ICML 2026
- SymmetricDiffusers: Learning Discrete Diffusion on Finite Symmetric GroupsYongxing Zhang, Donglin Yang, Renjie LiaoICLR 2025
它引用的顶会 Paper8
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- PolyGen: An Autoregressive Generative Model of 3D MeshesCharlie Nash, Yaroslav Ganin, S. M. Ali Eslami, Peter W. BattagliaICML 2020 · 被引用 339 次
- Fast Differentiable Sorting and RankingMathieu Blondel, Olivier Teboul, Quentin Berthet, Josip DjolongaICML 2020 · 被引用 285 次
相关 Paper
- Differentiable Sorting Networks for Scalable Sorting and Ranking SupervisionFelix Petersen, Christian Borgelt, Hilde Kuehne, Oliver DeussenICML 2021 · 被引用 39 次
- Monotonic Differentiable Sorting NetworksFelix Petersen, Christian Borgelt, Hilde Kuehne, Oliver DeussenICLR 2022 · 被引用 32 次
- OPS: An Order-Preserving Sorting Network for Information RetrievalChao Wang, Yongxiang Tang, Guikai Luan, Kaiyuan Li 等SIGIR 2026
- Divergence-Free Neural Networks with Application to Image DenoisingSébastien Herbreteau, Etienne MeunierICLR 2026
- LapSum - One Method to Differentiate Them All: Ranking, Sorting and Top-k SelectionLukasz Struski, Michal B. Bednarczyk, Igor T. Podolak, Jacek TaborICML 2025
