EEL: Efficiently Encoding Lattices for Reranking
Prasann Singhal, Jiacheng Xu, Xi Ye, Greg Durrett
摘要
Standard decoding approaches for conditional text generation tasks typically search for an output hypothesis with high model probability, but this may not yield the best hypothesis according to human judgments of quality. Reranking to optimize for “downstream” metrics can more closely optimize for quality, but many metrics of interest are computed with pre-trained language models, which are slow to apply to large numbers of hypotheses. We explore an approach for reranking hypotheses by using Transformers to efficiently encode lattices of generated outputs, a method we call EEL. With a single Transformer pass over the entire lattice, we can approximately compute a contextualized representation of each token as if it were only part of a single hypothesis in isolation. We combine this approach with a new class of token-factored rerankers (TFRs) that allow for efficient extraction of high reranker-scoring hypotheses from the lattice. Empirically, our approach incurs minimal degradation error compared to the exponentially slower approach of encoding each hypothesis individually. When applying EEL with TFRs across three text generation tasks, our results show both substantial speedup compared to naive reranking and often better performance on downstream metrics than comparable approaches.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- QUEST: Quality-Aware Metropolis-Hastings Sampling for Machine TranslationGonçalo Rui Alves Faria, Sweta Agrawal, António Farinhas, Ricardo Rei 等NeurIPS 2024 · 被引用 23 次
- Reranking Laws for Language Generation: A Communication-Theoretic PerspectiveAntónio Farinhas, Haau-Sing Li, André F. T. MartinsNeurIPS 2024 · 被引用 6 次
它引用的顶会 Paper13
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes 等ICLR 2020 · 被引用 4,112 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Plug and Play Language Models: A Simple Approach to Controlled Text GenerationSumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung 等ICLR 2020 · 被引用 1,166 次
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary 等ACL 2020 · 被引用 539 次
相关 Paper
- Discriminative Reranking for Neural Machine TranslationAnn Lee, Michael Auli, Marc'Aurelio RanzatoACL 2021
- Language Ranker: A Lightweight Ranking framework for LLM DecodingChenheng Zhang, Tianqi Du, Jizhe Zhang, Mingqing Xiao 等NeurIPS 2025 · 被引用 3 次
- Fast and Accurate Deep Bidirectional Language Representations for Unsupervised LearningJoongbo Shin, Yoonhyung Lee, Seunghyun Yoon, Kyomin JungACL 2020 · 被引用 6 次
- The Cascade Transformer: an Application for Efficient Answer Sentence SelectionLuca Soldaini, Alessandro MoschittiACL 2020 · 被引用 3 次
- ELMER: A Non-Autoregressive Pre-trained Language Model for Efficient and Effective Text GenerationJunyi Li, Tianyi Tang, Wayne Xin Zhao, Jian-Yun Nie 等EMNLP 2022 · 被引用 11 次
