Edit Flows: Variable Length Discrete Flow Matching with Sequence-Level Edit Operations
Marton Havasi, Brian Karrer, Itai Gat, Ricky T. Q. Chen
Abstract
Autoregressive generative models naturally generate variable-length sequences, while non-autoregressive models struggle, often imposing rigid, token-wise structures. We propose Edit Flows, a non-autoregressive model that overcomes these limitations by defining a discrete flow over sequences through edit operations— insertions, deletions, and substitutions. By modeling these operations within a Continuous-time Markov Chain over the sequence space, Edit Flows enable flexible, position-relative generation that aligns more closely with the structure of sequence data. Our training method leverages an expanded state space with auxiliary variables, making the learning process efficient and tractable. Empirical results show that Edit Flows outperforms both autoregressive and mask models on image captioning and significantly outperforms the mask construction in text and code generation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 073ff48b-6a59-4e74-aa94-883f6fd591b2Cited by top-tier papers4
- Flowception: Temporally Expansive Flow Matching for Video GenerationTariq Berrada Ifriqi, John Nguyen, Karteek Alahari, Jakob Verbeek et al.CVPR 2026 · 2 citations
- Insertion Based Sequence Generation with Learnable Order DynamicsDhruvesh Patel, Benjamin Rozonoyer, Gaurav Pandey, Tahira Naseem et al.ICML 2026 · 1 citation
- A Time-Reparameterized Cumulative Intensity Extrapolation Sampler for Discrete Flow MatchingFeiyang Fu, Hehe FanICML 2026
- STRIDE: Post-Training LLMs to Reason and Refine Bio-Sequences via Edit TrajectoriesDaiheng Zhang, Shiyang Zhang, Sizhuang He, Yangtian Zhang et al.ICML 2026
Builds on29
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari et al.ICML 2024 · 3,620 citations
- Structured Denoising Diffusion Models in Discrete State-SpacesJacob Austin, Daniel D. Johnson, Jonathan Ho, Daniel Tarlow et al.NeurIPS 2021 · 2,256 citations
Related papers
- Edit-Based Flow Matching for Temporal Point ProcessesDavid Lüdke, Marten Lienen, Marcel Kollovieh, Stephan GünnemannICLR 2026 · 9 citations
- Non-autoregressive Text Editing with Copy-aware Latent AlignmentsYu Zhang, Yue Zhang, Leyang Cui, Guohong FuEMNLP 2023 · 2 citations
- Autoregressive Image Generation without Vector QuantizationTianhong Li, Yonglong Tian, He Li, Mingyang Deng et al.NeurIPS 2024 · 758 citations
- MaskINT: Video Editing via Interpolative Non-autoregressive Masked TransformersHaoyu Ma, Shahin Mahdizadehaghdam, Bichen Wu, Zhipeng Fan et al.CVPR 2024 · 3 citations
- NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at ScaleChunrui Han, Guopeng Li, Jingwei Wu, Quan Sun et al.ICLR 2026 · 58 citations
