Lune

NeurIPS2020Top-tier venue

Multimodal Generative Learning Utilizing Jensen-Shannon-Divergence

Thomas M. Sutter, Imant Daunhawer, Julia E. Vogt

2020Year
105Citations
23Top-tier citations

Abstract

Learning from different data types is a long-standing goal in machine learning research, as multiple information sources co-occur when describing natural phenomena. However, existing generative models that approximate a multimodal ELBO rely on difficult or inefficient training schemes to learn a joint distribution and the dependencies between modalities. In this work, we propose a novel, efficient objective function that utilizes the Jensen-Shannon divergence for multiple distributions. It simultaneously approximates the unimodal and joint multimodal posteriors directly via a dynamic prior. In addition, we theoretically prove that the new multimodal JS-divergence (mmJSD) objective optimizes an ELBO. In extensive experiments, we demonstrate the advantage of the proposed mmJSD model compared to previous work in unsupervised, generative learning tasks. Joint and Conditional Generation. [27] implemented a multimodal VAE and introduced the idea that the distribution of the unimodal approximation should be close to the multimodal approximation function. [31] introduced the triple ELBO as an additional improvement. Both define labels as second modality and are not scalable in the number of modalities.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext de7e22d9-a6f4-4494-931e-0a8b270d3dd3

Cited by top-tier papers23

Ask how each one uses it

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines