Adaptive teachers for amortized samplers
Minsu Kim, Sanghyeok Choi, Taeyoung Yun, Emmanuel Bengio, Leo Feng, Jarrid Rector-Brooks, Sungsoo Ahn, Jinkyoo Park, Nikolay Malkin, Yoshua Bengio
摘要
Amortized inference is the task of training a parametric model, such as a neural network, to approximate a distribution with a given unnormalized density where exact sampling is intractable. When sampling is implemented as a sequential decision-making process, reinforcement learning (RL) methods, such as generative flow networks, can be used to train the sampling policy. Off-policy RL training facilitates the discovery of diverse, high-reward candidates, but existing methods still face challenges in efficient exploration. We propose to use an adaptive training distribution (the ) to guide the training of the primary amortized sampler (the ). The , an auxiliary behavior model, is trained to sample high-loss regions of the and can generalize across unexplored modes, thereby enhancing mode coverage by providing an efficient training curriculum. We validate the effectiveness of this approach in a synthetic environment designed to present an exploration challenge, two diffusion-based sampling tasks, and four biochemical discovery tasks demonstrating its ability to improve sample efficiency and mode coverage. Source code is available at https://github.com/alstn12088/adaptive-teacher.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Trust Region Constrained Measure Transport in Path Space for Stochastic Optimal Control and InferenceDenis Blessing, Julius Berner, Lorenz Richter, Carles Domingo-Enrich 等NeurIPS 2025 · 被引用 24 次
- On scalable and efficient training of diffusion samplersMinkyu Kim, Kiyoung Seong, Dongyeop Woo, Sungsoo Ahn 等NeurIPS 2025 · 被引用 11 次
- Bridge Matching Sampler: Scalable Sampling via Generalized Fixed-Point Diffusion MatchingDenis Blessing, Lorenz Richter, Julius Berner, Egor Malitskiy 等ICML 2026 · 被引用 8 次
- Active Attacks: Red-teaming LLMs via Adaptive EnvironmentsTaeyoung Yun, Pierre-Luc St-Charles, Jinkyoo Park, Yoshua Bengio 等ICML 2026 · 被引用 5 次
- Loss-Guided Auxiliary Agents for Overcoming Mode Collapse in GFlowNetsIdriss Malek, Aya Laajil, Abhijith Sharma, Eric Moulines 等AAAI 2026 · 被引用 3 次
它引用的顶会 Paper30
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford 等ICLR 2020 · 被引用 974 次
- Maximum Likelihood Training of Score-Based Diffusion ModelsYang Song, Conor Durkan, Iain Murray, Stefano ErmonNeurIPS 2021 · 被引用 958 次
- Flow Network based Generative Models for Non-Iterative Diverse Candidate GenerationEmmanuel Bengio, Moksh Jain, Maksym Korablyov, Doina Precup 等NeurIPS 2021 · 被引用 565 次
- Biological Sequence Design with GFlowNetsMoksh Jain, Emmanuel Bengio, Alex Hernández-García, Jarrid Rector-Brooks 等ICML 2022 · 被引用 224 次
相关 Paper
- Pre-Training and Fine-Tuning Generative Flow NetworksLing Pan, Moksh Jain, Kanika Madan, Yoshua BengioICLR 2024 · 被引用 24 次
- Improved off-policy training of diffusion samplersMarcin Sendera, Minsu Kim, Sarthak Mittal, Pablo Lemos 等NeurIPS 2024 · 被引用 52 次
- Local Search GFlowNetsMinsu Kim, Taeyoung Yun, Emmanuel Bengio, Dinghuai Zhang 等ICLR 2024 · 被引用 59 次
- Reinforced Sequential Monte Carlo for Amortised SamplingSanghyeok Choi, Sarthak Mittal, Víctor Elvira, Jinkyoo Park 等ICML 2026 · 被引用 2 次
- Avoid What You Know: Divergent Trajectory Balance for GFlowNetsPedro Dall’Antonia, Tiago Silva, Daniel Csillag, Salem Lahlou 等ICML 2026 · 被引用 2 次
