Goal-directed Generation of Discrete Structures with Conditional Generative Models
Amina Mollaysa, Brooks Paige, Alexandros Kalousis
摘要
Despite recent advances, goal-directed generation of structured discrete data remains challenging. For problems such as program synthesis (generating source code) and materials design (generating molecules), finding examples which satisfy desired constraints or exhibit desired properties is difficult. In practice, expensive heuristic search or reinforcement learning algorithms are often employed. In this paper we investigate the use of conditional generative models which directly attack this inverse problem, by modeling the distribution of discrete structures given properties of interest. Unfortunately, maximum likelihood training of such models often fails with the samples from the generative model inadequately respecting the input properties. To address this, we introduce a novel approach to directly optimize a reinforcement learning objective, maximizing an expected reward. We avoid high-variance score-function estimators that would otherwise be required by sampling from an approximation to the normalized rewards, allowing simple Monte Carlo estimation of model gradients. We test our methodology on two tasks: generating molecules with user-defined properties and identifying short python expressions which evaluate to a given target value. In both cases, we find improvements over maximum likelihood estimation and other baselines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Graph Diffusion Transformers for Multi-Conditional Molecular GenerationGang Liu, Jiaxin Xu, Tengfei Luo, Meng JiangNeurIPS 2024 · 被引用 73 次
- Learning Flexible Forward Trajectories for Masked Molecular DiffusionHyunjin Seo, Taewon Kim, Sihyun Yu, Sungsoo AhnICLR 2026 · 被引用 6 次
相关 Paper
- Feedback Efficient Online Fine-Tuning of Diffusion ModelsMasatoshi Uehara, Yulai Zhao, Kevin Black, Ehsan Hajiramezanali 等ICML 2024 · 被引用 47 次
- Fine-Tuning Discrete Diffusion Models via Reward Optimization with Applications to DNA and Protein DesignChenyu Wang, Masatoshi Uehara, Yichun He, Amy Wang 等ICLR 2025
- Reinforced Molecular Optimization with Neighborhood-Controlled GrammarsChencheng Xu, Qiao Liu, Minlie Huang, Tao JiangNeurIPS 2020 · 被引用 23 次
- Improving Molecular Design by Stochastic Iterative Target AugmentationKevin Yang, Wengong Jin, Kyle Swanson, Regina Barzilay 等ICML 2020 · 被引用 31 次
- Open Materials Generation with Inference-Time Reinforcement LearningPhilipp Höllmer, Stefano MartinianiICML 2026 · 被引用 3 次
