WR-One2Set: Towards Well-Calibrated Keyphrase Generation
Binbin Xie, Xiangpeng Wei, Baosong Yang, Huan Lin, Jun Xie, Xiaoli Wang, Min Zhang, Jinsong Su
摘要
Keyphrase generation aims to automatically generate short phrases summarizing an input document. The recently emerged ONE2SET paradigm (Ye et al., 2021) generates keyphrases as a set and has achieved competitive performance. Nevertheless, we observe serious calibration errors outputted by ONE2SET, especially in the over-estimation of ∅ token (means "no corresponding keyphrase"). In this paper, we deeply analyze this limitation and identify two main reasons behind: 1) the parallel generation has to introduce excessive ∅ as padding tokens into training instances; and 2) the training mechanism assigning target to each slot is unstable and further aggravates the ∅ token over-estimation. To make the model well-calibrated, we propose WR-ONE2SET which extends ONE2SET with an adaptive instance-level cost Weighting strategy and a target Re-assignment mechanism. The former dynamically penalizes the over-estimated slots for different instances thus smoothing the uneven training distribution. The latter refines the original inappropriate assignment and reduces the supervisory signals of over-estimated slots. Experimental results on commonly-used datasets demonstrate the effectiveness and generality of our proposed paradigm.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Towards Better Multi-modal Keyphrase Generation via Visual Entity Enhancement and Multi-granularity Image Noise FilteringYifan Dong, Suhang Wu, Fandong Meng, Jie Zhou 等ACM MM 2023 · 被引用 3 次
- One2Set + Large Language Model: Best Partners for Keyphrase GenerationLiangying Shao, Liang Zhang, Minlong Peng, Guoqi Ma 等EMNLP 2024 · 被引用 2 次
- Augmenting Intra-Modal Understanding in MLLMs for Robust Multimodal Keyphrase GenerationJiajun Cao, Qinggang Zhang, Yunbo Tang, Zhishang Xiang 等AAAI 2026
- Are Key-Phrases All That Reviewers Care About? A Comprehensive Benchmarking of Reviewer Matchmaking SystemsSourish Dasgupta, Harsh Sharma, Devansh Patel, Prarthee Desai 等AAAI 2025
它引用的顶会 Paper5
- One Size Does Not Fit All: Generating and Evaluating Variable Number of KeyphrasesXingdi Yuan, Tong Wang, Rui Meng, Khushboo Thaker 等ACL 2020 · 被引用 76 次
- Exclusive Hierarchical Decoding for Deep Keyphrase GenerationWang Chen, Hou Pong Chan, Piji Li, Irwin KingACL 2020 · 被引用 62 次
- A Variational Hierarchical Model for Neural Cross-Lingual SummarizationYunlong Liang, Fandong Meng, Chulun Zhou, Jinan Xu 等ACL 2022 · 被引用 36 次
- Fast and Constrained Absent Keyphrase Generation by Prompt-Based LearningHuanqin Wu, Baijiaxin Ma, Wei Liu, Tao Chen 等AAAI 2022 · 被引用 31 次
- One2Set: Generating Diverse Keyphrases as a SetJiacheng Ye, Tao Gui, Yichao Luo, Yige Xu 等ACL 2021
相关 Paper
- Keyphrase Generation via Soft and Hard Semantic CorrectionsGuangzhen Zhao, Guoshun Yin, Peng Yang, Yu YaoEMNLP 2022 · 被引用 3 次
- A Branching Decoder for Set GenerationZixian Huang, Gengyang Xiao, Yu Gu, Gong ChengICLR 2024 · 被引用 2 次
- Unsupervised Deep Keyphrase GenerationXianjie Shen, Yinghan Wang, Rui Meng, Jingbo ShangAAAI 2022 · 被引用 19 次
- Adaptive Beam Search Decoding for Discrete Keyphrase GenerationXiaoli Huang, Tongge Xu, Lvan Jiao, Yueran Zu 等AAAI 2021 · 被引用 10 次
- Unsupervised Open-domain Keyphrase GenerationLam Do, Pritom Saha Akash, Kevin Chen-Chuan ChangACL 2023 · 被引用 2 次
