Polynomial Width is Sufficient for Set Representation with High-dimensional Features
Peihao Wang, Shenghao Yang, Shu Li, Zhangyang Wang, Pan Li
摘要
Set representation has become ubiquitous in deep learning for modeling the inductive bias of neural networks that are insensitive to the input order. DeepSets is the most widely used neural network architecture for set representation. It involves embedding each set element into a latent space with dimension , followed by a sum pooling to obtain a whole-set embedding, and finally mapping the whole-set embedding to the output. In this work, we investigate the impact of the dimension on the expressive power of DeepSets. Previous analyses either oversimplified high-dimensional features to be one-dimensional features or were limited to analytic activations, thereby diverging from practical use or resulting in that grows exponentially with the set size and feature dimension . To investigate the minimal value of that achieves sufficient expressive power, we present two set-element embedding layers: (a) linear + power activation (LP) and (b) linear + exponential activations (LE). We demonstrate that being poly is sufficient for set representation using both embedding layers. We also provide a lower bound of for the LP embedding layer. Furthermore, we extend our results to permutation-equivariant set functions and the complex field.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Neural Injective Functions for Multisets, Measures and Graphs via a Finite Witness TheoremTal Amir, Steven J. Gortler, Ilai Avni, Ravina Ravina 等NeurIPS 2023 · 被引用 44 次
- Monotone and Separable Set Functions: Characterizations and Neural ModelsSoutrik Sarangi, Yonatan Sverdlov, Nadav Dym, Abir DeNeurIPS 2025 · 被引用 2 次
- Adversarial Encoding Perturbation and Synthesis for Set Representation Auxiliary LearningYankai Chen, Xinni Zhang, Henry Peng Zou, Bowei He 等ICLR 2026
- On the Hölder Stability of Multiset and Graph Neural NetworksYair Davidson, Nadav DymICLR 2025
- Latent Dimension Suffices for Universal Approximation of Permutation-invariant FunctionMin ZHOU, Enming Liang, Minghua ChenICML 2026
它引用的顶会 Paper11
- Principal Neighbourhood Aggregation for Graph NetsGabriele Corso, Luca Cavalleri, Dominique Beaini, Pietro Liò 等NeurIPS 2020 · 被引用 914 次
- Can Graph Neural Networks Count Substructures?Zhengdao Chen, Lei Chen, Soledad Villar, Joan BrunaNeurIPS 2020 · 被引用 392 次
- What graph neural networks cannot learn: depth vs widthAndreas LoukasICLR 2020 · 被引用 336 次
- Lorentz Group Equivariant Neural Network for Particle PhysicsAlexander Bogatskiy, Brandon M. Anderson, Jan T. Offermann, Marwah Roussi 等ICML 2020 · 被引用 164 次
- FSPool: Learning Set Representations with Featurewise Sort PoolingYan Zhang, Jonathon S. Hare, Adam Prügel-BennettICLR 2020 · 被引用 92 次
相关 Paper
- On the Representation Power of Set Pooling NetworksChristian Bueno, Alan HyltonNeurIPS 2021 · 被引用 13 次
- Exponential Separations in Symmetric Neural NetworksAaron Zweig, Joan BrunaNeurIPS 2022 · 被引用 10 次
- On Learning Sets of Symmetric ElementsHaggai Maron, Or Litany, Gal Chechik, Ethan FetayaICML 2020 · 被引用 148 次
- Stacking Deep Set Networks and Pooling by QuantilesZhuojun Chen, Xinghua Zhu, Dongzhe Su, Justin C. I. ChuangICML 2024 · 被引用 2 次
- On Universal Equivariant Set NetworksNimrod Segol, Yaron LipmanICLR 2020 · 被引用 74 次
