Differentiable Expectation-Maximization for Set Representation Learning
Minyoung Kim
摘要
We tackle the set2vec problem, the task of extracting a vector representation from an input set comprised of a variable number of feature vectors. Although recent approaches based on self attention such as (Set)Transformers were very successful due to the capability of capturing complex interaction between set elements, the computational overhead is the well-known downside. The inducing-point attention and the latest optimal transport kernel embedding (OTKE) are promising remedies that attain comparable or better performance with reduced computational cost, by incorporating a fixed number of learnable queries in attention. In this paper we approach the set2vec problem from a completely different perspective. The elements of an input set are considered as i.i.d. samples from a mixture distribution, and we define our set embedding feed-forward network as the maximum-a-posterior (MAP) estimate of the mixture which is approximately attained by a few Expectation-Maximization (EM) steps. The whole MAP-EM steps are differentiable operations with a fixed number of mixture parameters, allowing efficient auto-diff back-propagation for any given downstream task. Furthermore, the proposed mixture set data fitting framework allows unsupervised set representation learning naturally via marginal likelihood maximization aka the empirical Bayes. Interestingly, we also find that OTKE can be seen as a special case of our framework, specifically a single-step EM with extra balanced assignment constraints on the E-step. Compared to OTKE, our approach provides more flexible set embedding as well as prior-induced model regularization. We evaluate our approach on various tasks demonstrating improved performance over the state-of-the-arts.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper13
- Multimodal Prototyping for cancer survival predictionAndrew H. Song, Richard J. Chen, Guillaume Jaume, Anurag J. Vaidya 等ICML 2024 · 被引用 53 次
- Morphological Prototyping for Unsupervised Slide Representation Learning in Computational PathologyAndrew H. Song, Richard J. Chen, Tong Ding, Drew F. K. Williamson 等CVPR 2024 · 被引用 51 次
- DPsurv: Dual-Prototype Evidential Fusion for Uncertainty-Aware and Interpretable Whole Slide Image Survival PredictionYucheng Xing, ling huang, Jingying Ma, Ruping Hong 等ICML 2026 · 被引用 8 次
- Fisher Information Embedding for Node and Graph LearningDexiong Chen, Paolo Pellizzoni, Karsten M. BorgwardtICML 2023 · 被引用 4 次
- Scalable Set Encoding with Universal Mini-Batch Consistency and Unbiased Full Set Gradient ApproximationJeffrey Willette, Seanie Lee, Bruno Andreis, Kenji Kawaguchi 等ICML 2023 · 被引用 4 次
相关 Paper
- SeTformer Is What You Need for Vision and LanguagePourya Shamsolmoali, Masoumeh Zareapoor, Eric Granger, Michael FelsbergAAAI 2024 · 被引用 8 次
- A Trainable Optimal Transport Embedding for Feature Aggregation and its Relationship to AttentionGrégoire Mialon, Dexiong Chen, Alexandre d'Aspremont, Julien MairalICLR 2021 · 被引用 71 次
- Statistical Optimal Transport posed as Learning Kernel EmbeddingJagarlapudi Saketha Nath, Pratik Kumar JawanpuriaNeurIPS 2020 · 被引用 18 次
- Learning Prototype-oriented Set Representations for Meta-LearningDandan Guo, Long Tian, Minghe Zhang, Mingyuan Zhou 等ICLR 2022 · 被引用 27 次
- A VAE for Transformers with Nonparametric Variational Information BottleneckJames Henderson, Fabio FehrICLR 2023
