Multimodal Adversarially Learned Inference with Factorized Discriminators
Wenxue Chen, Jianke Zhu
Abstract
Learning from multimodal data is an important research topic in machine learning, which has the potential to obtain better representations. In this work, we propose a novel approach to generative modeling of multimodal data based on generative adversarial networks. To learn a coherent multimodal generative model, we show that it is necessary to align different encoder distributions with the joint decoder distribution simultaneously. To this end, we construct a specific form of the discriminator to enable our model to utilize data efficiently, which can be trained constrastively. By taking advantage of contrastive learning through factorizing the discriminator, we train our model on unimodal data. We have conducted experiments on the benchmark datasets, whose promising results show that our proposed approach outperforms the-state-ofthe-art methods on a variety of metrics. The source code is publicly available at https://github.com/6b5d/mmali.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0a07f09d-a127-447f-a691-29cdab43df74Cited by top-tier papers1
Ask how each one uses itBuilds on5
- Generalized Multimodal ELBOThomas M. Sutter, Imant Daunhawer, Julia E. VogtICLR 2021 · 130 citations
- Multimodal Generative Learning Utilizing Jensen-Shannon-DivergenceThomas M. Sutter, Imant Daunhawer, Julia E. VogtNeurIPS 2020 · 105 citations
- Relating by Contrasting: A Data-efficient Framework for Multimodal Generative ModelsYuge Shi, Brooks Paige, Philip H. S. Torr, N. SiddharthICLR 2021 · 42 citations
- Generalized Adversarially Learned InferenceYatin Dandi, Homanga Bharadhwaj, Abhishek Kumar, Piyush RaiAAAI 2021 · 8 citations
- Training Generative Adversarial Networks from Incomplete Observations using Factorised DiscriminatorsDaniel Stoller, Sebastian Ewert, Simon DixonICLR 2020 · 5 citations
Related papers
- Gaussian Mixture Variational Autoencoder with Contrastive Learning for Multi-Label ClassificationJunwen Bai, Shufeng Kong, Carla P. GomesICML 2022 · 48 citations
- Geometric Multimodal Contrastive Representation LearningPetra Poklukar, Miguel Vasco, Hang Yin, Francisco S. Melo et al.ICML 2022 · 68 citations
- Facilitating Multimodal Classification via Dynamically Learning Modality GapYang Yang, Fengqiang Wan, Qing-Yuan Jiang, Yi XuNeurIPS 2024 · 65 citations
- Multimodal Gaussian Mixture Variational Autoencoder with Consistency RegularizationsYarui Chen, Lehan Hong, Jianlin Shao, Jianning Yang et al.AAAI 2026
- Generative Modeling of Class Probability for Multi-Modal Representation LearningJungkyoo Shin, Bumsoo Kim, Eunwoo KimCVPR 2025
