GFlowNet-EM for Learning Compositional Latent Variable Models
Edward J. Hu, Nikolay Malkin, Moksh Jain, Katie E. Everett, Alexandros Graikos, Yoshua Bengio
摘要
Latent variable models (LVMs) with discrete compositional latents are an important but challenging setting due to a combinatorially large number of possible configurations of the latents. A key tradeoff in modeling the posteriors over latents is between expressivity and tractable optimization. For algorithms based on expectation-maximization (EM), the E-step is often intractable without restrictive approximations to the posterior. We propose the use of GFlowNets, algorithms for sampling from an unnormalized density by learning a stochastic policy for sequential construction of samples, for this intractable E-step. By training GFlowNets to sample from the posterior over latents, we take advantage of their strengths as amortized variational inference algorithms for complex distributions over discrete structures. Our approach, GFlowNet-EM, enables the training of expressive LVMs with discrete compositional latents, as shown by experiments on non-context-free grammar induction and on images using discrete variational autoencoders (VAEs) without conditional independence enforced in the encoder.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper28
- Learning GFlowNets From Partial Episodes For Improved Convergence And StabilityKanika Madan, Jarrid Rector-Brooks, Maksym Korablyov, Emmanuel Bengio 等ICML 2023 · 被引用 138 次
- Better Training of GFlowNets with Local Credit and Incomplete TrajectoriesLing Pan, Nikolay Malkin, Dinghuai Zhang, Yoshua BengioICML 2023 · 被引用 100 次
- Let the Flows Tell: Solving Graph Combinatorial Problems with GFlowNetsDinghuai Zhang, Hanjun Dai, Nikolay Malkin, Aaron C. Courville 等NeurIPS 2023 · 被引用 94 次
- Amortizing intractable inference in large language modelsEdward J. Hu, Moksh Jain, Eric Elmoznino, Younesse Kaddar 等ICLR 2024 · 被引用 91 次
- Joint Bayesian Inference of Graphical Structure and Parameters with a Single Generative Flow NetworkTristan Deleu, Mizu Nishikawa-Toomey, Jithendaraa Subramanian, Nikolay Malkin 等NeurIPS 2023 · 被引用 66 次
它引用的顶会 Paper12
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 被引用 9,451 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- Flow Network based Generative Models for Non-Iterative Diverse Candidate GenerationEmmanuel Bengio, Moksh Jain, Maksym Korablyov, Doina Precup 等NeurIPS 2021 · 被引用 565 次
- Learning GFlowNets From Partial Episodes For Improved Convergence And StabilityKanika Madan, Jarrid Rector-Brooks, Maksym Korablyov, Emmanuel Bengio 等ICML 2023 · 被引用 138 次
- Generative Flow Networks for Discrete Probabilistic ModelingDinghuai Zhang, Nikolay Malkin, Zhen Liu, Alexandra Volokhova 等ICML 2022 · 被引用 131 次
相关 Paper
- A theory of continuous generative flow networksSalem Lahlou, Tristan Deleu, Pablo Lemos, Dinghuai Zhang 等ICML 2023 · 被引用 118 次
- Streaming Bayes GFlowNetsTiago da Silva, Daniel Augusto de Souza, Diego MesquitaNeurIPS 2024 · 被引用 7 次
- Avoid What You Know: Divergent Trajectory Balance for GFlowNetsPedro Dall’Antonia, Tiago Silva, Daniel Csillag, Salem Lahlou 等ICML 2026 · 被引用 2 次
- Latent Logic Tree Extraction for Event Sequence Explanation from LLMsZitao Song, Chao Yang, Chaojie Wang, Bo An 等ICML 2024 · 被引用 11 次
- GFlowNets and variational inferenceNikolay Malkin, Salem Lahlou, Tristan Deleu, Xu Ji 等ICLR 2023
