LDMol: A Text-to-Molecule Diffusion Model with Structurally Informative Latent Space Surpasses AR Models
Jinho Chang, Jong Chul Ye
摘要
With the emergence of diffusion models as a frontline generative model, many researchers have proposed molecule generation techniques with conditional diffusion models. However, the unavoidable discreteness of a molecule makes it difficult for a diffusion model to connect raw data with highly complex conditions like natural language. To address this, here we present a novel latent diffusion model dubbed LDMol for text-conditioned molecule generation. By recognizing that the suitable latent space design is the key to the diffusion model performance, we employ a contrastive learning strategy to extract novel feature space from text data that embeds the unique characteristics of the molecule structure. Experiments show that LDMol outperforms the existing autoregressive baselines on the text-to-molecule generation benchmark, being one of the first diffusion models that outperforms autoregressive models in textual data generation with a better choice of the latent domain. Furthermore, we show that LDMol can be applied to downstream tasks such as moleculeto-text retrieval and text-guided molecule editing, demonstrating its versatility as a diffusion model.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- mCLM: A Modular Chemical Language Model that Generates Functional and Makeable MoleculesCarl Edwards, Chi Han, Gawon Lee, Thao Nguyen 等ICLR 2026 · 被引用 9 次
- Divide-and-Denoise: A Game-Theoretic Method for Fairly Composing Diffusion ModelsAbhi Gupta, Polina Barabanshchikova, Vikas Garg, Samuel Kaski 等ICML 2026
- BiMol-Diff: A Unified Diffusion Framework for Molecular Generation and CaptioningAditya Hemant Shahane, Anuj Kumar Sirohi, Devansh Arora, Nitin Kumar 等ACL 2026
- Controllable Molecule Generation via Sparse Representation Editing: An Interpretability-Driven PerspectiveZhuoran Li, Xu Sun, Chang Chen, Wanyu LINICML 2026
它引用的顶会 Paper28
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- Text-Guided Molecule Generation with Diffusion Language ModelHaisong Gong, Qiang Liu, Shu Wu, Liang WangAAAI 2024 · 被引用 45 次
- Geometric Latent Diffusion Models for 3D Molecule GenerationMinkai Xu, Alexander S. Powers, Ron O. Dror, Stefano Ermon 等ICML 2023 · 被引用 252 次
- Latent Diffusion for Language GenerationJustin Lovelace, Varsha Kishore, Chao Wan, Eliot Shekhtman 等NeurIPS 2023 · 被引用 177 次
- DisCo-Diff: Enhancing Continuous Diffusion Models with Discrete LatentsYilun Xu, Gabriele Corso, Tommi S. Jaakkola, Arash Vahdat 等ICML 2024 · 被引用 22 次
- NExT-Mol: 3D Diffusion Meets 1D Language Modeling for 3D Molecule GenerationZhiyuan Liu, Yanchen Luo, Han Huang, Enzhi Zhang 等ICLR 2025
