Tractable Control for Autoregressive Language Generation
Honghua Zhang, Meihua Dang, Nanyun Peng, Guy Van den Broeck
摘要
Despite the success of autoregressive large language models in text generation, it remains a major challenge to generate text that satisfies complex constraints: sampling from the conditional distribution is intractable for even the simplest lexical constraints . To overcome this challenge, we propose to use tractable probabilistic models (TPMs) to impose lexical constraints in autoregressive text generation models, which we refer to as GeLaTo (Generating Language with Tractable Constraints). To demonstrate the effectiveness of this framework, we use distilled hidden Markov models, where we can efficiently compute , to guide autoregressive generation from GPT2. GeLaTo achieves state-of-the-art performance on challenging benchmarks for constrained text generation (e.g., CommonGen), beating various strong baselines by a large margin. Our work not only opens up new avenues for controlling large language models but also motivates the development of more expressive TPMs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- Theoretical Benefit and Limitation of Diffusion Language ModelGuhao Feng, Yihan Geng, Jian Guan, Wei Wu 等NeurIPS 2025 · 被引用 52 次
- Scaling Tractable Probabilistic Circuits: A Systems PerspectiveAnji Liu, Kareem Ahmed, Guy Van den BroeckICML 2024 · 被引用 26 次
- Tailoring Self-Rationalizers with Multi-Reward DistillationSahana Ramnath, Brihi Joshi, Skyler Hallinan, Ximing Lu 等ICLR 2024 · 被引用 23 次
- On the Relationship Between Monotone and Squared Probabilistic CircuitsBenjie Wang, Guy Van den BroeckAAAI 2025 · 被引用 16 次
- Constrained Sampling for Language Models Should Be Easy: An MCMC PerspectiveEmmanuel Anaya Gonzalez, Sairam Vaidya, Kanghee Park, Ruyi Ji 等NeurIPS 2025 · 被引用 15 次
它引用的顶会 Paper13
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Einsum Networks: Fast and Scalable Learning of Tractable Probabilistic CircuitsRobert Peharz, Steven Lang, Antonio Vergari, Karl Stelzner 等ICML 2020 · 被引用 155 次
- Semantic Probabilistic Layers for Neuro-Symbolic LearningKareem Ahmed, Stefano Teso, Kai-Wei Chang, Guy Van den Broeck 等NeurIPS 2022 · 被引用 133 次
- POINTER: Constrained Progressive Text Generation via Insertion-based Generative Pre-trainingYizhe Zhang, Guoyin Wang, Chunyuan Li, Zhe Gan 等EMNLP 2020 · 被引用 68 次
- Controllable Text Generation with Neurally-Decomposed OracleTao Meng, Sidi Lu, Nanyun Peng, Kai-Wei ChangNeurIPS 2022 · 被引用 45 次
相关 Paper
- Control Large Language Models via Divide and ConquerBingxuan Li, Yiwei Wang, Tao Meng, Kai-Wei Chang 等EMNLP 2024 · 被引用 1 次
- Adaptable Logical Control for Large Language ModelsHonghua Zhang, Po-Nien Kung, Masahiro Yoshida, Guy Van den Broeck 等NeurIPS 2024 · 被引用 42 次
- Guiding LLMs The Right Way: Fast, Non-Invasive Constrained GenerationLuca Beurer-Kellner, Marc Fischer, Martin T. VechevICML 2024 · 被引用 93 次
- Guaranteed Generation from Large Language ModelsMinbeom Kim, Thibaut Thonet, Jos Rozen, Hwaran Lee 等ICLR 2025
- Syntactic and Semantic Control of Large Language Models via Sequential Monte CarloJoão Loula, Benjamin LeBrun, Li Du, Ben Lipkin 等ICLR 2025
