Multiset-Equivariant Set Prediction with Approximate Implicit Differentiation
Yan Zhang, David W. Zhang, Simon Lacoste-Julien, Gertjan J. Burghouts, Cees G. M. Snoek
Abstract
Most set prediction models in deep learning use set-equivariant operations, but they actually operate on multisets. We show that set-equivariant functions cannot represent certain functions on multisets, so we introduce the more appropriate notion of multiset-equivariance. We identify that the existing Deep Set Prediction Network (DSPN) can be multiset-equivariant without being hindered by set-equivariance and improve it with approximate implicit differentiation, allowing for better optimization while being faster and saving memory. In a range of toy experiments, we show that the perspective of multiset-equivariance is beneficial and that our changes to DSPN achieve better results in most cases. On CLEVR object property prediction, we substantially improve over the state-of-the-art Slot Attention from 8% to 77% in one of the strictest evaluation metrics because of the benefits made possible by implicit differentiation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- Equivariance with Learned Canonicalization FunctionsSékou-Oumar Kaba, Arnab Kumar Mondal, Yan Zhang, Yoshua Bengio et al.ICML 2023 · 109 citations
- Object Representations as Fixed Points: Training Iterative Refinement Algorithms with Implicit DifferentiationMichael Chang, Tom Griffiths, Sergey LevineNeurIPS 2022 · 69 citations
- Equivariant Networks for Crystal StructuresSékou-Oumar Kaba, Siamak RavanbakhshNeurIPS 2022 · 38 citations
- A Canonicalization Perspective on Invariant and Equivariant LearningGeorge Ma, Yifei Wang, Derek Lim, Stefanie Jegelka et al.NeurIPS 2024 · 38 citations
- Object centric architectures enable efficient causal representation learningAmin Mansouri, Jason S. Hartford, Yan Zhang, Yoshua BengioICLR 2024 · 28 citations
Builds on13
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran et al.NeurIPS 2020 · 1,275 citations
- Efficient and Modular Implicit DifferentiationMathieu Blondel, Quentin Berthet, Marco Cuturi, Roy Frostig et al.NeurIPS 2022 · 386 citations
- Multiscale Deep Equilibrium ModelsShaojie Bai, Vladlen Koltun, J. Zico KolterNeurIPS 2020 · 272 citations
- DeepSVG: A Hierarchical Generative Network for Vector Graphics AnimationAlexandre Carlier, Martin Danelljan, Alexandre Alahi, Radu TimofteNeurIPS 2020 · 247 citations
- Implicit Graph Neural NetworksFangda Gu, Heng Chang, Wenwu Zhu, Somayeh Sojoudi et al.NeurIPS 2020 · 188 citations
Related papers
- Invariant Slot Attention: Object Discovery with Slot-Centric Reference FramesOndrej Biza, Sjoerd van Steenkiste, Mehdi S. M. Sajjadi, Gamaleldin Fathy Elsayed et al.ICML 2023 · 54 citations
- On Learning Sets of Symmetric ElementsHaggai Maron, Or Litany, Gal Chechik, Ethan FetayaICML 2020 · 148 citations
- Stacking Deep Set Networks and Pooling by QuantilesZhuojun Chen, Xinghua Zhu, Dongzhe Su, Justin C. I. ChuangICML 2024 · 2 citations
- Unlocking Slot Attention by Changing Optimal Transport CostsYan Zhang, David W. Zhang, Simon Lacoste-Julien, Gertjan J. Burghouts et al.ICML 2023 · 20 citations
- Polynomial Width is Sufficient for Set Representation with High-dimensional FeaturesPeihao Wang, Shenghao Yang, Shu Li, Zhangyang Wang et al.ICLR 2024 · 9 citations
