Generative Marginalization Models
Sulin Liu, Peter J. Ramadge, Ryan P. Adams
摘要
We introduce marginalization models (MAMs), a new family of generative models for high-dimensional discrete data. They offer scalable and flexible generative modeling by explicitly modeling all induced marginal distributions. Marginalization models enable fast approximation of arbitrary marginal probabilities with a single forward pass of the neural network, which overcomes a major limitation of arbitrary marginal inference models, such as any-order autoregressive models. MAMs also address the scalability bottleneck encountered in training any-order generative models for high-dimensional problems under the context of energy-based training, where the goal is to match the learned distribution to a given desired probability (specified by an unnormalized log-probability function such as energy or reward function). We propose scalable methods for learning the marginals, grounded in the concept of"marginalization self-consistency". We demonstrate the effectiveness of the proposed model on a variety of discrete data distributions, including images, text, physical systems, and molecules, for maximum likelihood and energy-based training settings. MAMs achieve orders of magnitude speedup in evaluating the marginal probabilities on both settings. For energy-based training tasks, MAMs enable any-order generative modeling of high-dimensional problems beyond the scale of previous methods. Code is available at https://github.com/PrincetonLIPS/MaM.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Efficient Autoregressive Inference for Transformer Probabilistic ModelsConor Hassan, Nasrulloh R. B. S. Loka, Cen-You Li, Daolang Huang 等ICLR 2026 · 被引用 5 次
- Path-dependent Discrete Amortized InferenceTiago Silva, Esmeralda S. Whitammer, Salem LahlouICML 2026
- MetaDNS: Enhancing Exploration in Discrete Neural Samplers via MetadynamicsXiaochen Du, Juno Nam, Jaemoo Choi, Wei Guo 等ICML 2026
- The Limits of Tractable MarginalizationOliver Broadrick, Sanyam Agarwal, Guy Van den Broeck, Markus BläserICML 2025
它引用的顶会 Paper5
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Argmax Flows and Multinomial Diffusion: Learning Categorical DistributionsEmiel Hoogeboom, Didrik Nielsen, Priyank Jaini, Patrick Forré 等NeurIPS 2021 · 被引用 782 次
- Generative Flow Networks for Discrete Probabilistic ModelingDinghuai Zhang, Nikolay Malkin, Zhen Liu, Alexandra Volokhova 等ICML 2022 · 被引用 131 次
- Oops I Took A Gradient: Scalable Sampling for Discrete DistributionsWill Grathwohl, Kevin Swersky, Milad Hashemi, David Duvenaud 等ICML 2021 · 被引用 113 次
- Autoregressive Diffusion ModelsEmiel Hoogeboom, Alexey A. Gritsenko, Jasmijn Bastings, Ben Poole 等ICLR 2022 · 被引用 19 次
相关 Paper
- Generalized Energy Based ModelsMichael Arbel, Liang Zhou, Arthur GrettonICLR 2021 · 被引用 254 次
- Learning-Order Autoregressive Models with Application to Molecular Graph GenerationZhe Wang, Jiaxin Shi, Nicolas Heess, Arthur Gretton 等ICML 2025
- Joint Learning of Energy-based Models and their Partition FunctionMichael Eli Sander, Vincent Roulet, Tianlin Liu, Mathieu BlondelICML 2025
- Learning Energy-Based Models by Diffusion Recovery LikelihoodRuiqi Gao, Yang Song, Ben Poole, Ying Nian Wu 等ICLR 2021 · 被引用 144 次
- Semi-Autoregressive Energy Flows: Exploring Likelihood-Free Training of Normalizing FlowsPhillip Si, Zeyi Chen, Subham Sekhar Sahoo, Yair Schiff 等ICML 2023 · 被引用 9 次
