Incorporating Multimodal Information in Open-Domain Web Keyphrase Extraction
Yansen Wang, Zhen Fan, Carolyn P. Rosé
Abstract
Open-domain Keyphrase extraction (KPE) on the Web is a fundamental yet complex NLP task with a wide range of practical applications within the field of Information Retrieval. In contrast to other document types, web page designs are intended for easy navigation and information finding. Effective designs encode within the layout and formatting signals that point to where the important information can be found. In this work, we propose a modeling approach that leverages these multi-modal signals to aid in the KPE task. In particular, we leverage both lexical and visual features (e.g., size, font, position) at the micro-level to enable effective strategy induction, and metalevel features that describe pages at a macrolevel to aid in strategy selection. Our evaluation demonstrates that a combination of effective strategy induction and strategy selection within this approach for the KPE task outperforms state-of-the-art models. A qualitative post-hoc analysis illustrates how these features function within the model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Importance Estimation from Multiple Perspectives for Keyphrase ExtractionMingyang Song, Liping Jing, Lin XiaoEMNLP 2021 · 17 citations
- MUSTIE: Multimodal Structural Transformer for Web Information ExtractionQifan Wang, Jingang Wang, Xiaojun Quan, Fuli Feng et al.ACL 2023 · 16 citations
- RendNet: Unified 2D/3D Recognizer with Latent Space RenderingRuoxi Shi, Xinyang Jiang, Caihua Shan, Yansen Wang et al.CVPR 2022 · 4 citations
Related papers
- HyperRank: Hyperbolic Ranking Model for Unsupervised Keyphrase ExtractionMingyang Song, Huafeng Liu, Liping JingEMNLP 2023 · 5 citations
- Abstractive Open Information ExtractionKevin Pei, Ishan Jindal, Kevin Chen-Chuan ChangEMNLP 2023
- ZeroShotCeres: Zero-Shot Relation Extraction from Semi-Structured WebpagesColin Lockard, Prashant Shiralkar, Xin Luna Dong, Hannaneh HajishirziACL 2020 · 2 citations
- WIERT: Web Information Extraction via Render TreeZimeng Li, Bo Shao, Linjun Shou, Ming Gong et al.AAAI 2023 · 10 citations
- HiKEY: Hierarchical Multimodal Retrieval for Open-Domain Document Question AnsweringJoongmin Shin, Gyuho Shim, Jeongbae Park, Jaehyung Seo et al.ACL 2026
