Looking Backward: Retrospective Backward Synthesis for Goal-Conditioned GFlowNets
Haoran He, Can Chang, Huazhe Xu, Ling Pan
摘要
Generative Flow Networks (GFlowNets), a new family of probabilistic samplers, have demonstrated remarkable capabilities to generate diverse sets of high-reward candidates, in contrast to standard return maximization approaches (e.g., reinforcement learning) which often converge to a single optimal solution. Recent works have focused on developing goal-conditioned GFlowNets, which aim to train a single GFlowNet capable of achieving different outcomes as the task specifies. However, training such models is challenging due to extremely sparse rewards, particularly in high-dimensional problems. Moreover, previous methods suffer from the limited coverage of explored trajectories during training, which presents more pronounced challenges when only offline data is available. In this work, we propose a novel method called Retrospective Backward Synthesis (RBS) to address these critical problems. Specifically, RBS synthesizes new backward trajectories in goal-conditioned GFlowNets to enrich training trajectories with enhanced quality and diversity, thereby introducing copious learnable signals for effectively tackling the sparse reward problem. Extensive empirical results show that our method improves sample efficiency by a large margin and outperforms strong baselines on various standard evaluation benchmarks. Our codes are available at https://github.com/tinnerhrhe/Goal-Conditioned-GFN .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- FlowRL: Matching Reward Distributions for LLM ReasoningXuekai Zhu, Daixuan Cheng, Dinghuai Zhang, Hengli Li 等ICLR 2026 · 被引用 41 次
- Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet TrainingXi Wang, Wenbo Lu, Shenji WanICML 2026 · 被引用 1 次
- Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal ExamplesFangxu Yu, Lai Jiang, Haoqiang Kang, Shibo Hao 等ICML 2025
- Optimizing Backward Policies in GFlowNets via Trajectory Likelihood MaximizationTimofei Gritsaev, Nikita Morozov, Sergey Samsonov, Daniil TiapkinICLR 2025
- Pretraining Generative Flow Networks with Inexpensive Rewards for Molecular Graph GenerationMohit Pandey, Gopeshh Subbaraj, Artem Cherkasov, Martin Ester 等ICML 2025
它引用的顶会 Paper25
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 被引用 911 次
- Flow Network based Generative Models for Non-Iterative Diverse Candidate GenerationEmmanuel Bengio, Moksh Jain, Maksym Korablyov, Doina Precup 等NeurIPS 2021 · 被引用 565 次
- Contrastive Learning as Goal-Conditioned Reinforcement LearningBenjamin Eysenbach, Tianjun Zhang, Sergey Levine, Ruslan SalakhutdinovNeurIPS 2022 · 被引用 331 次
- Trajectory balance: Improved credit assignment in GFlowNetsNikolay Malkin, Moksh Jain, Emmanuel Bengio, Chen Sun 等NeurIPS 2022 · 被引用 316 次
- Biological Sequence Design with GFlowNetsMoksh Jain, Emmanuel Bengio, Alex Hernández-García, Jarrid Rector-Brooks 等ICML 2022 · 被引用 224 次
相关 Paper
- Pessimistic Backward Policy for GFlowNetsHyosoon Jang, Yunhui Jang, Minsu Kim, Jinkyoo Park 等NeurIPS 2024 · 被引用 14 次
- Local Search GFlowNetsMinsu Kim, Taeyoung Yun, Emmanuel Bengio, Dinghuai Zhang 等ICLR 2024 · 被引用 59 次
- GFlowNet Training by Policy GradientsPuhua Niu, Shili Wu, Mingzhou Fan, Xiaoning QianICML 2024 · 被引用 6 次
- Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet TrainingRuishuo Chen, Xun Wang, Rui Hu, Zhuoran Li 等ICML 2026
- Flow Factorization for Efficient Generative Flow NetworksJiashun Liu, Chunhui Li, Cheng-Hao Liu, Dianbo Liu 等AAAI 2025
