Text-Guided Molecule Generation with Diffusion Language Model
Haisong Gong, Qiang Liu, Shu Wu, Liang Wang
摘要
Text-guided molecule generation is a task where molecules are generated to match specific textual descriptions. Recently, most existing SMILES-based molecule generation methods rely on an autoregressive architecture. In this work, we propose the Text-Guided Molecule Generation with Diffusion Language Model (TGM-DLM), a novel approach that leverages diffusion models to address the limitations of autoregressive methods. TGM-DLM updates token embeddings within the SMILES string collectively and iteratively, using a two-phase diffusion generation process. The first phase optimizes embeddings from random noise, guided by the text description, while the second phase corrects invalid SMILES strings to form valid molecular representations. We demonstrate that TGM-DLM outperforms MolT5-Base, an autoregressive model, without the need for additional data resources. Our findings underscore the remarkable effectiveness of TGM-DLM in generating coherent and precise molecules with specific properties, opening new avenues in drug discovery and related scientific domains. Code will be released at: https://github.com/Deno-V/tgm-dlm.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Controllable Sequence Editing for Biological and Clinical TrajectoriesMichelle M. Li, Kevin Li, Yasha Ektefaie, Ying Jin 等ICLR 2026
- CombiMOTS: Combinatorial Multi-Objective Tree Search for Dual-Target Molecule GenerationThibaud Southiratn, Bonil Koo, Yijingxiu Lu, Sun KimICML 2025
- LDMol: A Text-to-Molecule Diffusion Model with Structurally Informative Latent Space Surpasses AR ModelsJinho Chang, Jong Chul YeICML 2025
- BiMol-Diff: A Unified Diffusion Framework for Molecular Generation and CaptioningAditya Hemant Shahane, Anuj Kumar Sirohi, Devansh Arora, Nitin Kumar 等ACL 2026
- Controllable Molecule Generation via Sparse Representation Editing: An Interpretability-Driven PerspectiveZhuoran Li, Xu Sun, Chang Chen, Wanyu LINICML 2026
它引用的顶会 Paper8
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- DiffWave: A Versatile Diffusion Model for Audio SynthesisZhifeng Kong, Wei Ping, Jiaji Huang, Kexin Zhao 等ICLR 2021 · 被引用 1,902 次
- Diffusion-LM Improves Controllable Text GenerationXiang Lisa Li, John Thickstun, Ishaan Gulrajani, Percy Liang 等NeurIPS 2022 · 被引用 1,546 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- DecompDiff: Diffusion Models with Decomposed Priors for Structure-Based Drug DesignJiaqi Guan, Xiangxin Zhou, Yuwei Yang, Yu Bao 等ICML 2023 · 被引用 115 次
相关 Paper
- GenMol: A Drug Discovery Generalist with Discrete DiffusionSeul Lee, Karsten Kreis, Srimukh Prasad Veccham, Meng Liu 等ICML 2025
- NExT-Mol: 3D Diffusion Meets 1D Language Modeling for 3D Molecule GenerationZhiyuan Liu, Yanchen Luo, Han Huang, Enzhi Zhang 等ICLR 2025
- DiffTMR: Diffusion-based Hierarchical Alignment for Text-Molecule RetrievalChenxu Wang, Dong Zhou, Ting Liu, Jianghao Lin 等ACM MM 2025
- ReAlign: Text-to-Motion Generation via Step-Aware Reward-Guided AlignmentWanjiang Weng, Xiaofeng Tan, Junbo Wang, Guo-Sen Xie 等AAAI 2026 · 被引用 6 次
- MolDiff: Addressing the Atom-Bond Inconsistency Problem in 3D Molecule Diffusion GenerationXingang Peng, Jiaqi Guan, Qiang Liu, Jianzhu MaICML 2023 · 被引用 76 次
