Incorporating Multimodal Information in Open-Domain Web Keyphrase Extraction
Yansen Wang, Zhen Fan, Carolyn P. Rosé
摘要
Open-domain Keyphrase extraction (KPE) on the Web is a fundamental yet complex NLP task with a wide range of practical applications within the field of Information Retrieval. In contrast to other document types, web page designs are intended for easy navigation and information finding. Effective designs encode within the layout and formatting signals that point to where the important information can be found. In this work, we propose a modeling approach that leverages these multi-modal signals to aid in the KPE task. In particular, we leverage both lexical and visual features (e.g., size, font, position) at the micro-level to enable effective strategy induction, and metalevel features that describe pages at a macrolevel to aid in strategy selection. Our evaluation demonstrates that a combination of effective strategy induction and strategy selection within this approach for the KPE task outperforms state-of-the-art models. A qualitative post-hoc analysis illustrates how these features function within the model.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Importance Estimation from Multiple Perspectives for Keyphrase ExtractionMingyang Song, Liping Jing, Lin XiaoEMNLP 2021 · 被引用 17 次
- MUSTIE: Multimodal Structural Transformer for Web Information ExtractionQifan Wang, Jingang Wang, Xiaojun Quan, Fuli Feng 等ACL 2023 · 被引用 16 次
- RendNet: Unified 2D/3D Recognizer with Latent Space RenderingRuoxi Shi, Xinyang Jiang, Caihua Shan, Yansen Wang 等CVPR 2022 · 被引用 4 次
相关 Paper
- HyperRank: Hyperbolic Ranking Model for Unsupervised Keyphrase ExtractionMingyang Song, Huafeng Liu, Liping JingEMNLP 2023 · 被引用 5 次
- Abstractive Open Information ExtractionKevin Pei, Ishan Jindal, Kevin Chen-Chuan ChangEMNLP 2023
- ZeroShotCeres: Zero-Shot Relation Extraction from Semi-Structured WebpagesColin Lockard, Prashant Shiralkar, Xin Luna Dong, Hannaneh HajishirziACL 2020 · 被引用 2 次
- WIERT: Web Information Extraction via Render TreeZimeng Li, Bo Shao, Linjun Shou, Ming Gong 等AAAI 2023 · 被引用 10 次
- HiKEY: Hierarchical Multimodal Retrieval for Open-Domain Document Question AnsweringJoongmin Shin, Gyuho Shim, Jeongbae Park, Jaehyung Seo 等ACL 2026
