Generative property enhancer: implicit guided generation through conditional density estimation
Pedro O. Pinheiro, Pan Kessel, Aya Abdelsalam Ismail, Sai Pooja Mahajan, Kyunghyun Cho, Saeed Saremi, Natasa Tagasovska
Abstract
Generative modeling is increasingly important for data-driven computational design. Conventional approaches pair a generative model with a discriminative model to select or guide samples toward optimized designs. Yet discriminative models often struggle in data-scarce settings, common in scientific applications, and are unreliable in the tails of the distribution where optimal designs typically lie. We introduce generative property enhancer (GPE), an approach that implicitly guides generation by matching samples with lower property values to higher-value ones. Formulated as conditional density estimation, our framework defines a target distribution with improved properties, compelling the generative model to produce enhanced, diverse designs without auxiliary predictors. GPE is simple, scalable, end-to-end, modality-agnostic, and integrates seamlessly with diverse generative model architectures and losses. We demonstrate competitive empirical results on standard in silico offline (non-sequential) protein fitness optimization benchmarks. Finally, we propose iterative training on a combination of limited real data and self-generated synthetic data, enabling extrapolation beyond the original property ranges.
- This work was done while the author was at Genentech. 1 Here, "design optimization" is taken in its broader, applied meaning common in science and engineering, different from its formal mathematical definition.
39th Conference on Neural Information Processing Systems (NeurIPS 2025).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on37
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- Extracting Training Data from Large Language ModelsNicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski et al.USENIX Security 2021 · 2,866 citations
Related papers
- Implicitly Guided Design with PropEn: Match your Data to Follow the GradientNatasa Tagasovska, Vladimir Gligorijevic, Kyunghyun Cho, Andreas LoukasNeurIPS 2024 · 10 citations
- Improving Molecular Design by Stochastic Iterative Target AugmentationKevin Yang, Wengong Jin, Kyle Swanson, Regina Barzilay et al.ICML 2020 · 31 citations
- Aligning Protein Conformation Ensemble Generation with Physical FeedbackJiarui Lu, Xiaoyin Chen, Stephen Zhewen Lu, Aurélie C. Lozano et al.ICML 2025
- A Variational Perspective on Generative Protein Fitness OptimizationLea Bogensperger, Dominik Narnhofer, Ahmed Allam, Konrad Schindler et al.ICML 2025
- Robust Model-Based Optimization for Challenging Fitness LandscapesSaba Ghaffari, Ehsan Saleh, Alexander G. Schwing, Yu-Xiong Wang et al.ICLR 2024 · 3 citations
