Gradient-guided Unsupervised Lexically Constrained Text Generation
Lei Sha
Abstract
Lexically-constrained generation requires the target sentence to satisfy some lexical constraints, such as containing some specific words or being the paraphrase to a given sentence, which is very important in many real-world natural language generation applications. Previous works usually apply beamsearch-based methods or stochastic searching methods to lexically-constrained generation. However, when the search space is too large, beam-search-based methods always fail to find the constrained optimal solution. At the same time, stochastic search methods always cost too many steps to find the correct optimization direction. In this paper, we propose a novel method G2LC to solve the lexically-constrained generation as an unsupervised gradient-guided optimization problem. We propose a differentiable objective function and use the gradient to help determine which position in the sequence should be changed (deleted or inserted/replaced by another word). The word updating process of the inserted/replaced word also benefits from the guidance of gradient. Besides, our method is free of parallel data training, which is flexible to be used in the inference stage of any pre-trained generation model. We apply G2LC to two generation tasks: keyword-to-sentence generation and unsupervised paraphrase generation. The experiment results show that our method achieves state-of-the-art compared to previous lexically-constrained methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 20e93eff-0513-4d7c-9569-eafdc57c70b5Cited by top-tier papers21
- Diffusion-LM Improves Controllable Text GenerationXiang Lisa Li, John Thickstun, Ishaan Gulrajani, Percy Liang et al.NeurIPS 2022 · 1,546 citations
- COLD Decoding: Energy-based Constrained Text Generation with Langevin DynamicsLianhui Qin, Sean Welleck, Daniel Khashabi, Yejin ChoiNeurIPS 2022 · 217 citations
- Amortizing intractable inference in large language modelsEdward J. Hu, Moksh Jain, Eric Elmoznino, Younesse Kaddar et al.ICLR 2024 · 91 citations
- Learning from the Best: Rationalizing Predictions by Adversarial Information CalibrationLei Sha, Oana-Maria Camburu, Thomas LukasiewiczAAAI 2021 · 40 citations
- Parallel Refinements for Lexically Constrained Text Generation with BARTXingwei HeEMNLP 2021 · 34 citations
Builds on2
Related papers
- ParaLS: Lexical Substitution via Pretrained ParaphraserJipeng Qiang, Kang Liu, Yun Li, Yunhao Yuan et al.ACL 2023 · 9 citations
- Show Me How To Revise: Improving Lexically Constrained Sentence Generation with XLNetXingwei He, Victor O. K. LiAAAI 2021 · 26 citations
- Explicit Syntactic Guidance for Neural Text GenerationYafu Li, Leyang Cui, Jianhao Yan, Yongjing Yin et al.ACL 2023 · 4 citations
- POINTER: Constrained Progressive Text Generation via Insertion-based Generative Pre-trainingYizhe Zhang, Guoyin Wang, Chunyuan Li, Zhe Gan et al.EMNLP 2020 · 68 citations
- Control Large Language Models via Divide and ConquerBingxuan Li, Yiwei Wang, Tao Meng, Kai-Wei Chang et al.EMNLP 2024 · 1 citation
