Multi-Label Zero-Shot Product Attribute-Value Extraction
Jiaying Gong, Hoda Eldardiry
Abstract
E-commerce platforms should provide detailed product descriptions (attribute values) for effective product search and recommendation. However, attribute value information is typically not available for new products. To predict unseen attribute values, large quantities of labeled training data are needed to train a traditional supervised learning model. Typically, it is difficult, time-consuming, and costly to manually label large quantities of new product profiles. In this paper, we propose a novel method to efficiently and effectively extract unseen attribute values from new products in the absence of labeled data (zero-shot setting). We propose HyperPAVE, a multilabel zero-shot attribute value extraction model that leverages inductive inference in heterogeneous hypergraphs. In particular, our proposed technique constructs heterogeneous hypergraphs to capture complex higher-order relations (i.e. user behavior information) to learn more accurate feature representations for graph nodes. Furthermore, our proposed HyperPAVE model uses an inductive link prediction mechanism to infer future connections between unseen nodes. This enables HyperPAVE to identify new attribute values without the need for labeled training data. We conduct extensive experiments with ablation studies on different categories of the MAVE dataset. The results demonstrate that our proposed HyperPAVE model significantly outperforms existing classificationbased, generation-based large language models for attribute value extraction in the zero-shot setting. CCS CONCEPTS • Computing methodologies → Information extraction.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 12ce368f-695d-4379-8dbc-c7922dd0b353Cited by top-tier papers1
Ask how each one uses itBuilds on14
- Inductive representation learning on temporal graphsDa Xu, Chuanwei Ruan, Evren Körpeoglu, Sushant Kumar et al.ICLR 2020 · 901 citations
- Be More with Less: Hypergraph Attention Networks for Inductive Text ClassificationKaize Ding, Jianling Wang, Jundong Li, Dingcheng Li et al.EMNLP 2020 · 210 citations
- Label Verbalization and Entailment for Effective Zero and Few-Shot Relation ExtractionOscar Sainz, Oier Lopez de Lacalle, Gorka Labaka, Ander Barrena et al.EMNLP 2021 · 94 citations
- HGMF: Heterogeneous Graph-based Fusion for Multimodal Data with IncompletenessJiayi Chen, Aidong ZhangKDD 2020 · 89 citations
- Learning to Extract Attribute Value from Product via Question Answering: A Multi-task ApproachQifan Wang, Li Yang, Bhargav Kanagal, Sumit Sanghai et al.KDD 2020 · 75 citations
Related papers
- Generalize to Fully Unseen Graphs: Learn Transferable Hyper-Relation Structures for Inductive Link PredictionJing Yang, Xiaowen Jiang, Yuan Gao, Laurence T. Yang et al.ACM MM 2024 · 5 citations
- HYPER: A Foundation Model for Inductive Link Prediction with Knowledge HypergraphsXingyue Huang, Mikhail Galkin, Michael M. Bronstein, Ismail Ilkan CeylanICLR 2026 · 12 citations
- Multimodal Joint Attribute Prediction and Value Extraction for E-commerce ProductTiangang Zhu, Yue Wang, Haoran Li, Youzheng Wu et al.EMNLP 2020 · 46 citations
- Towards Multimodal Inductive Learning: Adaptively Embedding MMKG via PrototypesShundong Yang, Jing Yang, Xiaowen Jiang, Yuan Gao et al.WWW 2025 · 3 citations
- THGB: A Comprehensive Benchmark for Text-attributed Heterogeneous GraphsLixin Zhou, Zemin Liu, Yuan Fang, Dan Niu et al.AAAI 2026
