Learning to Scale Logits for Temperature-Conditional GFlowNets
Minsu Kim, Joohwan Ko, Taeyoung Yun, Dinghuai Zhang, Ling Pan, Woochang Kim, Jinkyoo Park, Emmanuel Bengio, Yoshua Bengio
摘要
GFlowNets are probabilistic models that sequentially generate compositional structures through a stochastic policy. Among GFlowNets, temperature-conditional GFlowNets can introduce temperature-based controllability for exploration and exploitation. We propose Logit-scaling GFlowNets (Logit-GFN), a novel architectural design that greatly accelerates the training of temperature-conditional GFlowNets. It is based on the idea that previously proposed approaches introduced numerical challenges in the deep network training, since different temperatures may give rise to very different gradient profiles as well as magnitudes of the policy's logits. We find that the challenge is greatly reduced if a learned function of the temperature is used to scale the policy's logits directly. Also, using Logit-GFN, GFlowNets can be improved by having better generalization capabilities in offline learning and mode discovery capabilities in online learning, which is empirically verified in various biological and chemical tasks. Our code is available at https://github.com/dbsxodud-11/logit-gfn
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Improved off-policy training of diffusion samplersMarcin Sendera, Minsu Kim, Sarthak Mittal, Pablo Lemos 等NeurIPS 2024 · 被引用 52 次
- Genetic-guided GFlowNets for Sample Efficient Molecular OptimizationHyeonah Kim, Minsu Kim, Sanghyeok Choi, Jinkyoo ParkNeurIPS 2024 · 被引用 42 次
- QGFN: Controllable Greediness with Action ValuesElaine Lau, Stephen Zhewen Lu, Ling Pan, Doina Precup 等NeurIPS 2024 · 被引用 21 次
- Order-Preserving GFlowNetsYihang Chen, Lukas MauchICLR 2024 · 被引用 17 次
- Pessimistic Backward Policy for GFlowNetsHyosoon Jang, Yunhui Jang, Minsu Kim, Jinkyoo Park 等NeurIPS 2024 · 被引用 14 次
它引用的顶会 Paper27
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar 等ICLR 2021 · 被引用 1,270 次
- Flow Network based Generative Models for Non-Iterative Diverse Candidate GenerationEmmanuel Bengio, Moksh Jain, Maksym Korablyov, Doina Precup 等NeurIPS 2021 · 被引用 565 次
- Trajectory balance: Improved credit assignment in GFlowNetsNikolay Malkin, Moksh Jain, Emmanuel Bengio, Chen Sun 等NeurIPS 2022 · 被引用 316 次
- Biological Sequence Design with GFlowNetsMoksh Jain, Emmanuel Bengio, Alex Hernández-García, Jarrid Rector-Brooks 等ICML 2022 · 被引用 224 次
相关 Paper
- Pre-Training and Fine-Tuning Generative Flow NetworksLing Pan, Moksh Jain, Kanika Madan, Yoshua BengioICLR 2024 · 被引用 24 次
- Loss-Guided Auxiliary Agents for Overcoming Mode Collapse in GFlowNetsIdriss Malek, Aya Laajil, Abhijith Sharma, Eric Moulines 等AAAI 2026 · 被引用 3 次
- Routing by Reaching: Composition of Pre-trained GFlowNets for Multi-Objective GenerationSeokwon Yoon, Youngbin Choi, Seunghyuk Cho, Seungbeom Lee 等ICML 2026
- Local Search GFlowNetsMinsu Kim, Taeyoung Yun, Emmanuel Bengio, Dinghuai Zhang 等ICLR 2024 · 被引用 59 次
- GFlowNet-EM for Learning Compositional Latent Variable ModelsEdward J. Hu, Nikolay Malkin, Moksh Jain, Katie E. Everett 等ICML 2023 · 被引用 48 次
