Text-Guided Molecule Generation with Diffusion Language Model
Haisong Gong, Qiang Liu, Shu Wu, Liang Wang
Abstract
Text-guided molecule generation is a task where molecules are generated to match specific textual descriptions. Recently, most existing SMILES-based molecule generation methods rely on an autoregressive architecture. In this work, we propose the Text-Guided Molecule Generation with Diffusion Language Model (TGM-DLM), a novel approach that leverages diffusion models to address the limitations of autoregressive methods. TGM-DLM updates token embeddings within the SMILES string collectively and iteratively, using a two-phase diffusion generation process. The first phase optimizes embeddings from random noise, guided by the text description, while the second phase corrects invalid SMILES strings to form valid molecular representations. We demonstrate that TGM-DLM outperforms MolT5-Base, an autoregressive model, without the need for additional data resources. Our findings underscore the remarkable effectiveness of TGM-DLM in generating coherent and precise molecules with specific properties, opening new avenues in drug discovery and related scientific domains. Code will be released at: https://github.com/Deno-V/tgm-dlm.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- Controllable Sequence Editing for Biological and Clinical TrajectoriesMichelle M. Li, Kevin Li, Yasha Ektefaie, Ying Jin et al.ICLR 2026
- CombiMOTS: Combinatorial Multi-Objective Tree Search for Dual-Target Molecule GenerationThibaud Southiratn, Bonil Koo, Yijingxiu Lu, Sun KimICML 2025
- LDMol: A Text-to-Molecule Diffusion Model with Structurally Informative Latent Space Surpasses AR ModelsJinho Chang, Jong Chul YeICML 2025
- BiMol-Diff: A Unified Diffusion Framework for Molecular Generation and CaptioningAditya Hemant Shahane, Anuj Kumar Sirohi, Devansh Arora, Nitin Kumar et al.ACL 2026
- Controllable Molecule Generation via Sparse Representation Editing: An Interpretability-Driven PerspectiveZhuoran Li, Xu Sun, Chang Chen, Wanyu LINICML 2026
Builds on8
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- DiffWave: A Versatile Diffusion Model for Audio SynthesisZhifeng Kong, Wei Ping, Jiaji Huang, Kexin Zhao et al.ICLR 2021 · 1,902 citations
- Diffusion-LM Improves Controllable Text GenerationXiang Lisa Li, John Thickstun, Ishaan Gulrajani, Percy Liang et al.NeurIPS 2022 · 1,546 citations
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- DecompDiff: Diffusion Models with Decomposed Priors for Structure-Based Drug DesignJiaqi Guan, Xiangxin Zhou, Yuwei Yang, Yu Bao et al.ICML 2023 · 115 citations
Related papers
- GenMol: A Drug Discovery Generalist with Discrete DiffusionSeul Lee, Karsten Kreis, Srimukh Prasad Veccham, Meng Liu et al.ICML 2025
- NExT-Mol: 3D Diffusion Meets 1D Language Modeling for 3D Molecule GenerationZhiyuan Liu, Yanchen Luo, Han Huang, Enzhi Zhang et al.ICLR 2025
- DiffTMR: Diffusion-based Hierarchical Alignment for Text-Molecule RetrievalChenxu Wang, Dong Zhou, Ting Liu, Jianghao Lin et al.ACM MM 2025
- ReAlign: Text-to-Motion Generation via Step-Aware Reward-Guided AlignmentWanjiang Weng, Xiaofeng Tan, Junbo Wang, Guo-Sen Xie et al.AAAI 2026 · 6 citations
- MolDiff: Addressing the Atom-Bond Inconsistency Problem in 3D Molecule Diffusion GenerationXingang Peng, Jiaqi Guan, Qiang Liu, Jianzhu MaICML 2023 · 76 citations
