Expected Information Maximization: Using the I-Projection for Mixture Density Estimation
Philipp Becker, Oleg Arenz, Gerhard Neumann
摘要
Modelling highly multi-modal data is a challenging problem in machine learning. Most algorithms are based on maximizing the likelihood, which corresponds to the M(oment)-projection of the data distribution to the model distribution. The M-projection forces the model to average over modes it cannot represent. In contrast, the I(information)-projection ignores such modes in the data and concentrates on the modes the model can represent. Such behavior is appealing whenever we deal with highly multi-modal data where modelling single modes correctly is more important than covering all the modes. Despite this advantage, the I-projection is rarely used in practice due to the lack of algorithms that can efficiently optimize it based on data. In this work, we present a new algorithm called Expected Information Maximization (EIM) for computing the I-projection solely based on samples for general latent variable models, where we focus on Gaussian mixtures models and Gaussian mixtures of experts. Our approach applies a variational upper bound to the I-projection objective which decomposes the original objective into single objectives for each mixture component as well as for the coefficients, allowing an efficient optimization. Similar to GANs, our approach employs discriminators but uses a more stable optimization procedure, using a tight upper bound. We show that our algorithm is much more effective in computing the I-projection than recent GAN approaches and we illustrate the effectiveness of our approach for modelling multi-modal behavior on two pedestrian and traffic prediction datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Adversarial Imitation Learning with PreferencesAleksandar Taranovic, Andras Gabor Kupcsik, Niklas Freymuth, Gerhard NeumannICLR 2023 · 被引用 25 次
- Variational Distillation of Diffusion Policies into Mixture of ExpertsHongyi Zhou, Denis Blessing, Ge Li, Onur Celik 等NeurIPS 2024 · 被引用 19 次
- Learning Mixtures of Experts with EM: A Mirror Descent PerspectiveQuentin Fruytier, Aryan Mokhtari, Sujay SanghaviICML 2025
相关 Paper
- Unity by Diversity: Improved Representation Learning for Multimodal VAEsThomas M. Sutter, Yang Meng, Andrea Agostini, Daphné Chopard 等NeurIPS 2024 · 被引用 21 次
- HMGAN: A Hierarchical Multi-Modal Generative Adversarial Network Model for Wearable Human Activity RecognitionLing Chen, Rong Hu, Menghan Wu, Xin ZhouUbiComp 2023 · 被引用 23 次
- Variational Pedestrian DetectionYuang Zhang, Huanyu He, Jianguo Li, Yuxi Li 等CVPR 2021
- Federated-EM with heterogeneity mitigation and variance reductionAymeric Dieuleveut, Gersende Fort, Eric Moulines, Geneviève RobinNeurIPS 2021 · 被引用 28 次
- A Stochastic Path Integral Differential EstimatoR Expectation Maximization AlgorithmGersende Fort, Eric Moulines, Hoi-To WaiNeurIPS 2020 · 被引用 9 次
