Evidential Softmax for Sparse Multimodal Distributions in Deep Generative Models
Phil Chen, Masha Itkina, Ransalu Senanayake, Mykel J. Kochenderfer
摘要
Many applications of generative models rely on the marginalization of their highdimensional output probability distributions. Normalization functions that yield sparse probability distributions can make exact marginalization more computationally tractable. However, sparse normalization functions usually require alternative loss functions for training since the log-likelihood is undefined for sparse probability distributions. Furthermore, many sparse normalization functions often collapse the multimodality of distributions. In this work, we present ev-softmax, a sparse normalization function that preserves the multimodality of probability distributions. We derive its properties, including its gradient in closed-form, and introduce a continuous family of approximations to ev-softmax that have full support and can be trained with probabilistic loss functions such as negative log-likelihood and Kullback-Leibler divergence. We evaluate our method on a variety of generative models, including variational autoencoders and auto-regressive architectures. Our method outperforms existing dense and sparse normalization techniques in distributional accuracy. We demonstrate that ev-softmax successfully reduces the dimensionality of probability distributions while maintaining multimodality.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Adaptive Compositional Continual Meta-LearningBin Wu, Jinyuan Fang, Xiangxiang Zeng, Shangsong Liang 等ICML 2023 · 被引用 12 次
- Failures Are Fated, But Can Be Faded: Characterizing and Mitigating Unwanted Behaviors in Large-Scale Vision and Language ModelsSom Sagar, Aditya Taparia, Ransalu SenanayakeICML 2024 · 被引用 11 次
- Learning Discrete Structured Variational Auto-Encoder using Natural Evolution StrategiesAlon Berliner, Guy Rotman, Yossi Adi, Roi Reichart 等ICLR 2022 · 被引用 5 次
- MultiMax: Sparse and Multi-Modal Attention LearningYuxuan Zhou, Mario Fritz, Margret KeuperICML 2024 · 被引用 4 次
- Random-Set Neural NetworksShireen Kudukkil Manchingal, Muhammad Mubashar, Kaizheng Wang, Keivan Shariatmadar 等ICLR 2025
它引用的顶会 Paper2
- Efficient Marginalization of Discrete and Structured Latent Variables via SparsityGonçalo M. Correia, Vlad Niculae, Wilker Aziz, André F. T. MartinsNeurIPS 2020 · 被引用 25 次
- Evidential Sparsification of Multimodal Latent Spaces in Conditional Variational AutoencodersMasha Itkina, Boris Ivanovic, Ransalu Senanayake, Mykel J. Kochenderfer 等NeurIPS 2020 · 被引用 21 次
相关 Paper
- SurVAE Flows: Surjections to Bridge the Gap between VAEs and FlowsDidrik Nielsen, Priyank Jaini, Emiel Hoogeboom, Ole Winther 等NeurIPS 2020 · 被引用 100 次
- A Batch Normalized Inference Network Keeps the KL Vanishing AwayQile Zhu, Wei Bi, Xiaojiang Liu, Xiyao Ma 等ACL 2020 · 被引用 70 次
- Sparse and Continuous Attention MechanismsAndré F. T. Martins, António Farinhas, Marcos V. Treviso, Vlad Niculae 等NeurIPS 2020 · 被引用 55 次
- Generative Marginalization ModelsSulin Liu, Peter J. Ramadge, Ryan P. AdamsICML 2024 · 被引用 3 次
- Improving Variational Autoencoders with Density Gap-based RegularizationJianfei Zhang, Jun Bai, Chenghua Lin, Yanmeng Wang 等NeurIPS 2022 · 被引用 11 次
