Parallel Refinements for Lexically Constrained Text Generation with BART
Xingwei He
Abstract
Lexically constrained text generation aims to control the generated text by incorporating some pre-specified keywords into the output. Previous work injects lexical constraints into the output by controlling the decoding process or refining the candidate output iteratively, which tends to generate generic or ungrammatical sentences, and has high computational complexity. To address these challenges, we propose Constrained BART (CBART) for lexically constrained text generation. CBART leverages the pre-trained model BART and transfers part of the generation burden from the decoder to the encoder by decomposing this task into two sub-tasks, thereby improving the sentence quality. Concretely, we extend BART by adding a token-level classifier over the encoder, aiming at instructing the decoder where to replace and insert. Guided by the encoder, the decoder refines multiple tokens of the input in one step by inserting tokens before specific positions and re-predicting tokens with low confidence. To further reduce the inference latency, the decoder predicts all tokens in parallel. Experiment results on One-Billion-Word and Yelp show that CBART can generate plausible text with high quality and diversity while significantly accelerating inference.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 70a740c5-18f8-42c6-b2e2-0b4340cf153dCited by top-tier papers10
- MeaCap: Memory-Augmented Zero-shot Image CaptioningZequn Zeng, Yan Xie, Hao Zhang, Chiyu Chen et al.CVPR 2024 · 38 citations
- UCEpic: Unifying Aspect Planning and Lexical Constraints for Generating Explanations in RecommendationJiacheng Li, Zhankui He, Jingbo Shang, Julian J. McAuleyKDD 2023 · 12 citations
- Flexible-length Text Infilling for Discrete Diffusion ModelsAndrew Zhang, Anushka Sivakumar, Chia-Wei Tang, Chris ThomasEMNLP 2025 · 11 citations
- Metric-guided Distillation: Distilling Knowledge from the Metric to Ranker and Retriever for Generative Commonsense ReasoningXingwei He, Yeyun Gong, A-Long Jin, Weizhen Qi et al.EMNLP 2022 · 10 citations
- Improving Factual Error Correction by Learning to Inject Factual ErrorsXingwei He, Qianru Zhang, A-Long Jin, Jun Ma et al.AAAI 2024 · 5 citations
Builds on6
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes et al.ICLR 2020 · 4,112 citations
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- POINTER: Constrained Progressive Text Generation via Insertion-based Generative Pre-trainingYizhe Zhang, Guoyin Wang, Chunyuan Li, Zhe Gan et al.EMNLP 2020 · 68 citations
- Gradient-guided Unsupervised Lexically Constrained Text GenerationLei ShaEMNLP 2020 · 35 citations
- Show Me How To Revise: Improving Lexically Constrained Sentence Generation with XLNetXingwei He, Victor O. K. LiAAAI 2021 · 26 citations
Related papers
- PAIR: Planning and Iterative Refinement in Pre-trained Transformers for Long Text GenerationXinyu Hua, Lu WangEMNLP 2020 · 44 citations
- Control Large Language Models via Divide and ConquerBingxuan Li, Yiwei Wang, Tao Meng, Kai-Wei Chang et al.EMNLP 2024 · 1 citation
- Neural Rule-Execution Tracking Machine For Transformer-Based Text GenerationYufei Wang, Can Xu, Huang Hu, Chongyang Tao et al.NeurIPS 2021 · 11 citations
- Generic resources are what you need: Style transfer tasks without task-specific parallel training dataHuiyuan Lai, Antonio Toral, Malvina NissimEMNLP 2021 · 14 citations
- Knowledge Infused DecodingRuibo Liu, Guoqing Zheng, Shashank Gupta, Radhika Gaonkar et al.ICLR 2022 · 18 citations
