DiffusER: Diffusion via Edit-based Reconstruction
Machel Reid, Vincent Josua Hellendoorn, Graham Neubig
Abstract
In text generation, models that generate text from scratch one token at a time are currently the dominant paradigm. Despite being performant, these models lack the ability to revise existing text, which limits their usability in many practical scenarios. We look to address this, with DiffusER (Diffusion via Edit-based Reconstruction), a new edit-based generative model for text based on denoising diffusion models -- a class of models that use a Markov chain of denoising steps to incrementally generate data. DiffusER is not only a strong generative model in general, rivalling autoregressive models on several tasks spanning machine translation, summarization, and style transfer; it can also perform other varieties of generation that standard autoregressive models are not well-suited for. For instance, we demonstrate that DiffusER makes it possible for a user to condition generation on a prototype, or an incomplete sequence, and continue revising based on previous edit steps.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cbc15e5e-7fe3-48f5-882f-97df59bc6f8cCited by top-tier papers10
- Aligning LLM Agents by Learning Latent Preference from User EditsGe Gao, Alexey Taymanov, Eduardo Salinas, Paul Mineiro et al.NeurIPS 2024 · 102 citations
- Fast Solvers for Discrete Diffusion Models: Theory and Applications of High-Order AlgorithmsYinuo Ren, Haoxuan Chen, Yuchen Zhu, Wei Guo et al.NeurIPS 2025 · 51 citations
- DreamOn: Diffusion Language Models For Code Infilling Beyond Fixed-size CanvasZirui Wu, Lin Zheng, Zhihui Xie, Jiacheng Ye et al.ICLR 2026 · 33 citations
- Edit Flows: Variable Length Discrete Flow Matching with Sequence-Level Edit OperationsMarton Havasi, Brian Karrer, Itai Gat, Ricky T. Q. ChenNeurIPS 2025 · 12 citations
- Promises and Pitfalls of Generative Masked Language Modeling: Theoretical Framework and Practical GuidelinesYuchen Li, Alexandre Kirchmeyer, Aashay Mehta, Yilong Qin et al.ICML 2024 · 5 citations
Builds on12
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes et al.ICLR 2020 · 4,112 citations
- Diffusion-LM Improves Controllable Text GenerationXiang Lisa Li, John Thickstun, Ishaan Gulrajani, Percy Liang et al.NeurIPS 2022 · 1,546 citations
- Cold Diffusion: Inverting Arbitrary Image Transforms Without NoiseArpit Bansal, Eitan Borgnia, Hong-Min Chu, Jie Li et al.NeurIPS 2023 · 469 citations
Related papers
- AR-Diffusion: Auto-Regressive Diffusion Model for Text GenerationTong Wu, Zhihao Fan, Xiao Liu, Hai-Tao Zheng et al.NeurIPS 2023 · 170 citations
- PLANNER: Generating Diversified Paragraph via Latent Language Diffusion ModelYizhe Zhang, Jiatao Gu, Zhuofeng Wu, Shuangfei Zhai et al.NeurIPS 2023 · 65 citations
- DiffuSeq: Sequence to Sequence Text Generation with Diffusion ModelsShansan Gong, Mukai Li, Jiangtao Feng, Zhiyong Wu et al.ICLR 2023 · 94 citations
- Flexible-length Text Infilling for Discrete Diffusion ModelsAndrew Zhang, Anushka Sivakumar, Chia-Wei Tang, Chris ThomasEMNLP 2025 · 11 citations
- Text Diffusion with Reinforced ConditioningYuxuan Liu, Tianchi Yang, Shaohan Huang, Zihan Zhang et al.AAAI 2024 · 2 citations
