Set Norm and Equivariant Skip Connections: Putting the Deep in Deep Sets
Lily H. Zhang, Veronica Tozzo, John M. Higgins, Rajesh Ranganath
摘要
Permutation invariant neural networks are a promising tool for making predictions from sets. However, we show that existing permutation invariant architectures, Deep Sets and Set Transformer, can suffer from vanishing or exploding gradients when they are deep. Additionally, layer norm, the normalization of choice in Set Transformer, can hurt performance by removing information useful for prediction. To address these issues, we introduce the "clean path principle" for equivariant residual connections and develop set norm (sn), a normalization tailored for sets. With these, we build Deep Sets++ and Set Transformer++, models that reach high depths with better or comparable performance than their original counterparts on a diverse suite of tasks. We additionally introduce Flow-RBC, a new single-cell dataset and real-world application of permutation invariant prediction. We open-source our data and code here: https://github.com/rajesh-lab/deep_permutation_invariant.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Robust Self-Supervised Multi-Instance Learning with Structure AwarenessYejiang Wang, Yuhai Zhao, Zhengkui Wang, Meixia WangAAAI 2023 · 被引用 6 次
- Improving Set Function Approximation with Quasi-Arithmetic Neural NetworksTomás Tokár, Scott SannerICLR 2026
- Breaking the Simplification Bottleneck in Amortized Neural Symbolic RegressionPaul Saegert, Ullrich KoetheICML 2026
- Generative Distribution Embeddings: Lifting autoencoders to the space of distributions for multiscale representation learningNic Fishman, Gokul Gowri, Peng Yin, Jonathan Gootenberg 等NeurIPS 2025
- Wasserstein Flow Matching: Generative Modeling Over Families of DistributionsDoron Haviv, Aram-Alexandre Pooladian, Dana Pe'er, Brandon AmosICML 2025
它引用的顶会 Paper8
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding 等ICML 2020 · 被引用 1,910 次
- DeepGCNs: Can GCNs Go As Deep As CNNs?Guohao Li, Matthias Müller, Ali K. Thabet, Bernard GhanemICCV 2019 · 被引用 1,586 次
- On Layer Normalization in the Transformer ArchitectureRuibin Xiong, Yunchang Yang, Di He, Kai Zheng 等ICML 2020 · 被引用 1,388 次
- PairNorm: Tackling Oversmoothing in GNNsLingxiao Zhao, Leman AkogluICLR 2020 · 被引用 590 次
- Revisiting Point Cloud Shape Classification with a Simple and Effective BaselineAnkit Goyal, Hei Law, Bowei Liu, Alejandro Newell 等ICML 2021 · 被引用 297 次
相关 Paper
- On Universal Equivariant Set NetworksNimrod Segol, Yaron LipmanICLR 2020 · 被引用 74 次
- Scalable Normalizing Flows for Permutation Invariant DensitiesMarin Bilos, Stephan GünnemannICML 2021 · 被引用 28 次
- DuMLP-Pin: A Dual-MLP-Dot-Product Permutation-Invariant Network for Set Feature ExtractionJiajun Fei, Ziyu Zhu, Wenlei Liu, Zhidong Deng 等AAAI 2022 · 被引用 6 次
- Exchangeable Neural ODE for Set ModelingYang Li, Haidong Yi, Christopher M. Bender, Siyuan Shan 等NeurIPS 2020 · 被引用 32 次
- On the Representation Power of Set Pooling NetworksChristian Bueno, Alan HyltonNeurIPS 2021 · 被引用 13 次
