On Divergence Measures for Training GFlowNets
Tiago da Silva, Eliezer de Souza da Silva, Diego Mesquita
Abstract
Generative Flow Networks (GFlowNets) are amortized inference models designed to sample from unnormalized distributions over composable objects, with applications in generative modeling for tasks in fields such as causal discovery, NLP, and drug discovery. Traditionally, the training procedure for GFlowNets seeks to minimize the expected log-squared difference between a proposal (forward policy) and a target (backward policy) distribution, which enforces certain flow-matching conditions. While this training procedure is closely related to variational inference (VI), directly attempting standard Kullback-Leibler (KL) divergence minimization can lead to proven biased and potentially high-variance estimators. Therefore, we first review four divergence measures, namely, Renyi-'s, Tsallis-'s, reverse and forward KL's, and design statistically efficient estimators for their stochastic gradients in the context of training GFlowNets. Then, we verify that properly minimizing these divergences yields a provably correct and empirically effective training scheme, often leading to significantly faster convergence than previously proposed optimization. To achieve this, we design control variates based on the REINFORCE leave-one-out and score-matching estimators to reduce the variance of the learning objectives'gradients. Our work contributes by narrowing the gap between GFlowNets training and generalized variational approximations, paving the way for algorithmic ideas informed by the divergence minimization viewpoint.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Evaluating GFlowNet from partial episodes for stable and flexible policy-based trainingPuhua Niu, Shili Wu, Xiaoning QianICLR 2026 · 2 citations
- -Trajectory Balance: A Loss Family for Tuning GFlowNets, Generative Models, and LLMs with Off- and On-Policy DataJake Fawkes, Jason HartfordICML 2026
- Revisiting Non-Acyclic GFlowNets in Discrete EnvironmentsNikita Morozov, Ian Maksimov, Daniil Tiapkin, Sergey SamsonovICML 2025
- Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet TrainingRuishuo Chen, Xun Wang, Rui Hu, Zhuoran Li et al.ICML 2026
Builds on26
- Flow Network based Generative Models for Non-Iterative Diverse Candidate GenerationEmmanuel Bengio, Moksh Jain, Maksym Korablyov, Doina Precup et al.NeurIPS 2021 · 565 citations
- Trajectory balance: Improved credit assignment in GFlowNetsNikolay Malkin, Moksh Jain, Emmanuel Bengio, Chen Sun et al.NeurIPS 2022 · 316 citations
- Biological Sequence Design with GFlowNetsMoksh Jain, Emmanuel Bengio, Alex Hernández-García, Jarrid Rector-Brooks et al.ICML 2022 · 224 citations
- Learning GFlowNets From Partial Episodes For Improved Convergence And StabilityKanika Madan, Jarrid Rector-Brooks, Maksym Korablyov, Emmanuel Bengio et al.ICML 2023 · 138 citations
- Generative Flow Networks for Discrete Probabilistic ModelingDinghuai Zhang, Nikolay Malkin, Zhen Liu, Alexandra Volokhova et al.ICML 2022 · 131 citations
Related papers
- GFlowNets and variational inferenceNikolay Malkin, Salem Lahlou, Tristan Deleu, Xu Ji et al.ICLR 2023
- A theory of continuous generative flow networksSalem Lahlou, Tristan Deleu, Pablo Lemos, Dinghuai Zhang et al.ICML 2023 · 118 citations
- Beyond Squared Error: Exploring Loss Design for Enhanced Training of Generative Flow NetworksRui Hu, Yifan Zhang, Zhuoran Li, Longbo HuangICLR 2025
- Avoid What You Know: Divergent Trajectory Balance for GFlowNetsPedro Dall’Antonia, Tiago Silva, Daniel Csillag, Salem Lahlou et al.ICML 2026 · 2 citations
- Optimizing Backward Policies in GFlowNets via Trajectory Likelihood MaximizationTimofei Gritsaev, Nikita Morozov, Sergey Samsonov, Daniil TiapkinICLR 2025
