Generative Modeling of Discrete Latent Structures via Dynamic Policy Gradients
Stefan Ivanovic, Ge Liu, Mohammed El-Kebir
Abstract
Many scientific problems require inferring unobserved mechanistic latent states from indirect observations. While classical approaches, including expectation maximization, do not scale to combinatorially large spaces, deep learning approaches such as variational autoencoders typically form artificial latent states rather than reconstructing the mechanistic ground-truth states. Here, we introduce GReinSS, a policy learning framework that uses dynamically rescaled rewards to learn latent state distributions that maximize the observed data likelihood. We show that GReinSS accurately reconstructs simulated latent sets and latent graphs, outperforming alternative policy learning and generative modeling baselines. Additionally, GReinSS reconstructs isoforms from real short-read RNA sequencing data that better match isoforms detected by orthogonal long-read sequencing than the standard RSEM algorithm. Overall, GReinSS is a principled and practically effective approach for generative modeling and inference of combinatorial latent states from indirect observations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 54b537a7-8d0d-4941-a1bb-1086bb59ed2bBuilds on3
- Structured Denoising Diffusion Models in Discrete State-SpacesJacob Austin, Daniel D. Johnson, Jonathan Ho, Daniel Tarlow et al.NeurIPS 2021 · 2,256 citations
- Trajectory balance: Improved credit assignment in GFlowNetsNikolay Malkin, Moksh Jain, Emmanuel Bengio, Chen Sun et al.NeurIPS 2022 · 316 citations
- Automatic Intrinsic Reward Shaping for Exploration in Deep Reinforcement LearningMingqi Yuan, Bo Li, Xin Jin, Wenjun ZengICML 2023 · 17 citations
Related papers
- RIDER: 3D RNA Inverse Design with Reinforcement Learning-Guided DiffusionTianmeng Hu, Yongzheng Cui, Biao Luo, Ke LiICLR 2026 · 1 citation
- Structure-based RNA Design by Step-wise Optimization of Latent Diffusion ModelQi Si, Xuyang Liu, Penglei Wang, Xin Guo et al.AAAI 2026
- GFlowNet-EM for Learning Compositional Latent Variable ModelsEdward J. Hu, Nikolay Malkin, Moksh Jain, Katie E. Everett et al.ICML 2023 · 48 citations
- LIMO: Latent Inceptionism for Targeted Molecule GenerationPeter Eckmann, Kunyang Sun, Bo Zhao, Mudong Feng et al.ICML 2022 · 60 citations
- Structured Flow Autoencoders: Learning Structured Probabilistic Representations with Flow MatchingYidan Xu, Yixin Wang, XuanLong NguyenICLR 2026
