Rethinking Model Selection and Decoding for Keyphrase Generation with Pre-trained Sequence-to-Sequence Models
Di Wu, Wasi Uddin Ahmad, Kai-Wei Chang
摘要
Keyphrase Generation (KPG) is a longstanding task in NLP with widespread applications. The advent of sequence-to-sequence (seq2seq) pre-trained language models (PLMs) has ushered in a transformative era for KPG, yielding promising performance improvements. However, many design decisions remain unexplored and are often made arbitrarily. This paper undertakes a systematic analysis of the influence of model selection and decoding strategies on PLM-based KPG. We begin by elucidating why seq2seq PLMs are apt for KPG, anchored by an attention-driven hypothesis. We then establish that conventional wisdom for selecting seq2seq PLMs lacks depth: (1) merely increasing model size or performing task-specific adaptation is not parameter-efficient; (2) although combining in-domain pre-training with task adaptation benefits KPG, it does partially hinder generalization. Regarding decoding, we demonstrate that while greedy search achieves strong F1 scores, it lags in recall compared with sampling-based methods. Based on these insights, we propose DeSel, a likelihood-based decode-select algorithm for seq2seq PLMs. DeSel improves greedy search by an average of 4.7% semantic F1 across five datasets. Our collective findings pave the way for deeper future investigations into PLM-based KPG.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- One2Set + Large Language Model: Best Partners for Keyphrase GenerationLiangying Shao, Liang Zhang, Minlong Peng, Guoqi Ma 等EMNLP 2024 · 被引用 2 次
- Multi-Task Knowledge Distillation with Embedding Constraints for Scholarly Keyphrase Boundary ClassificationSeo Park, Cornelia CarageaEMNLP 2023 · 被引用 1 次
它引用的顶会 Paper12
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Cross-Task Generalization via Natural Language Crowdsourcing InstructionsSwaroop Mishra, Daniel Khashabi, Chitta Baral, Hannaneh HajishirziACL 2022 · 被引用 887 次
- S2ORC: The Semantic Scholar Open Research CorpusKyle Lo, Lucy Lu Wang, Mark Neumann, Rodney Kinney 等ACL 2020 · 被引用 424 次
- Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP TasksYizhong Wang, Swaroop Mishra, Pegah Alipoormolabashi, Yeganeh Kordi 等EMNLP 2022 · 被引用 238 次
相关 Paper
- Adaptive Beam Search Decoding for Discrete Keyphrase GenerationXiaoli Huang, Tongge Xu, Lvan Jiao, Yueran Zu 等AAAI 2021 · 被引用 10 次
- MUDY: Multi-Granular Dynamic Candidate Contextualization for Unsupervised Keyphrase ExtractionHyeongu Kang, Susik YoonSIGIR 2026
- PromptRank: Unsupervised Keyphrase Extraction Using PromptAobo Kong, Shiwan Zhao, Hao Chen, Qicheng Li 等ACL 2023 · 被引用 31 次
- Fast and Constrained Absent Keyphrase Generation by Prompt-Based LearningHuanqin Wu, Baijiaxin Ma, Wei Liu, Tao Chen 等AAAI 2022 · 被引用 31 次
- RaSE-KGC: A Relation-Aware Segment Encoding Approach for Knowledge Graph CompletionChenxiao Lin, Ye Luo, Kunhong Liu, Qingqiang WuICDE 2026
