InsNet: An Efficient, Flexible, and Performant Insertion-based Text Generation Model
Sidi Lu, Tao Meng, Nanyun Peng
摘要
We propose INSNET, an expressive insertion-based text generator with efficient training and flexible decoding (parallel or sequential). Unlike most existing insertion-based text generation works that require re-encoding of the context after each insertion operation and thus are inefficient to train, INSNET only requires one pass of context encoding for the entire sequence during training by introducing a novel insertion-oriented position encoding and a light-weighted slot representation strategy to enable computation sharing. Furthermore, we propose an algorithm INSNET-Dinic to better determine the parallelization of insertion operations that provides a controllable switch between parallel and sequential decoding, making it flexible to handle more parallelizable tasks such as machine translation with efficient decoding, or less parallelizable tasks such as open-domain text generation to guarantee high-quality outputs. Experiments on two lexically constrained text generation datasets and three machine translation datasets demonstrate IN-SNET's advantages over previous insertion-based methods in terms of training speed, inference efficiency, and generation quality. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Amortizing intractable inference in large language modelsEdward J. Hu, Moksh Jain, Eric Elmoznino, Younesse Kaddar 等ICLR 2024 · 被引用 91 次
- Tractable Control for Autoregressive Language GenerationHonghua Zhang, Meihua Dang, Nanyun Peng, Guy Van den BroeckICML 2023 · 被引用 63 次
- AMOM: Adaptive Masking over Masking for Conditional Masked Language ModelYisheng Xiao, Ruiyang Xu, Lijun Wu, Juntao Li 等AAAI 2023 · 被引用 14 次
- Flexible-length Text Infilling for Discrete Diffusion ModelsAndrew Zhang, Anushka Sivakumar, Chia-Wei Tang, Chris ThomasEMNLP 2025 · 被引用 11 次
- Beyond Masks: Efficient, Flexible Diffusion Language Models via Deletion-Insertion ProcessesFangyu Ding, Ding Ding, Sijin Chen, Kaibo Wang 等ICLR 2026 · 被引用 7 次
它引用的顶会 Paper3
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Content Planning for Neural Story Generation with Aristotelian RescoringSeraphina Goldfarb-Tarrant, Tuhin Chakrabarty, Ralph M. Weischedel, Nanyun PengEMNLP 2020 · 被引用 106 次
- Context-Situated Pun GenerationJiao Sun, Anjali Narayan-Chen, Shereen Oraby, Shuyang Gao 等EMNLP 2022 · 被引用 7 次
相关 Paper
- RenewNAT: Renewing Potential Translation for Non-autoregressive TransformerPei Guo, Yisheng Xiao, Juntao Li, Min ZhangAAAI 2023 · 被引用 9 次
- TITE: Token-Independent Text Encoder for Information RetrievalFerdinand Schlatt, Tim Hagen, Martin Potthast, Matthias HagenSIGIR 2025 · 被引用 2 次
- Glancing Transformer for Non-Autoregressive Neural Machine TranslationLihua Qian, Hao Zhou, Yu Bao, Mingxuan Wang 等ACL 2021
- POINTER: Constrained Progressive Text Generation via Insertion-based Generative Pre-trainingYizhe Zhang, Guoyin Wang, Chunyuan Li, Zhe Gan 等EMNLP 2020 · 被引用 68 次
- InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-trainingDingdong Wang, Jin Xu, Ruihang Chu, Zhifang Guo 等ACL 2025 · 被引用 9 次
