ERU-KG: Efficient Reference-aligned Unsupervised Keyphrase Generation
Lam Thanh Do, Aaditya Bodke, Pritom Saha Akash, Kevin Chen-Chuan Chang
摘要
Unsupervised keyphrase prediction has gained growing interest in recent years. However, existing methods typically rely on heuristically defined importance scores, which may lead to inaccurate informativeness estimation. In addition, they lack consideration for time efficiency. To solve these problems, we propose ERU-KG, an unsupervised keyphrase generation (UKG) model that consists of an informativeness and a phraseness module. The former estimates the relevance of keyphrase candidates, while the latter generate those candidates. The informativeness module innovates by learning to model informativeness through references (e.g., queries, citation contexts, and titles) and at the term-level, thereby 1) capturing how the key concepts of documents are perceived in different contexts and 2) estimating informativeness of phrases more efficiently by aggregating term informativeness, removing the need for explicit modeling of the candidates. ERU-KG demonstrates its effectiveness on keyphrase generation benchmarks by outperforming unsupervised baselines and achieving on average 89% of the performance of a supervised model for top 10 predictions. Additionally, to highlight its practical utility, we evaluate the model on text retrieval tasks and show that keyphrases generated by ERU-KG are effective when employed as query and document expansions. Furthermore, inference speed tests reveal that ERU-KG is the fastest among baselines of similar model sizes. Finally, our proposed model can switch between keyphrase generation and extraction by adjusting hyperparameters, catering to diverse application requirements. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- AttentionRank: Unsupervised Keyphrase Extraction using Self and Cross AttentionsHaoran Ding, Xiao LuoEMNLP 2021 · 被引用 46 次
- SciRepEval: A Multi-Format Benchmark for Scientific Document RepresentationsAmanpreet Singh, Mike D'Arcy, Arman Cohan, Doug Downey 等EMNLP 2023 · 被引用 45 次
- PromptRank: Unsupervised Keyphrase Extraction Using PromptAobo Kong, Shiwan Zhao, Hao Chen, Qicheng Li 等ACL 2023 · 被引用 31 次
- SPECTER: Document-level Representation Learning using Citation-informed TransformersArman Cohan, Sergey Feldman, Iz Beltagy, Doug Downey 等ACL 2020 · 被引用 20 次
- Unsupervised Deep Keyphrase GenerationXianjie Shen, Yinghan Wang, Rui Meng, Jingbo ShangAAAI 2022 · 被引用 19 次
相关 Paper
- Unsupervised Open-domain Keyphrase GenerationLam Do, Pritom Saha Akash, Kevin Chen-Chuan ChangACL 2023 · 被引用 2 次
- HyperRank: Hyperbolic Ranking Model for Unsupervised Keyphrase ExtractionMingyang Song, Huafeng Liu, Liping JingEMNLP 2023 · 被引用 5 次
- Heterogeneous Graph Neural Networks for Keyphrase GenerationJiacheng Ye, Ruijian Cai, Tao Gui, Qi ZhangEMNLP 2021 · 被引用 14 次
- Fast and Constrained Absent Keyphrase Generation by Prompt-Based LearningHuanqin Wu, Baijiaxin Ma, Wei Liu, Tao Chen 等AAAI 2022 · 被引用 31 次
- Importance Estimation from Multiple Perspectives for Keyphrase ExtractionMingyang Song, Liping Jing, Lin XiaoEMNLP 2021 · 被引用 17 次
