Multimodal Joint Attribute Prediction and Value Extraction for E-commerce Product
Tiangang Zhu, Yue Wang, Haoran Li, Youzheng Wu, Xiaodong He, Bowen Zhou
摘要
Product attribute values are essential in many e-commerce scenarios, such as customer service robots, product recommendations, and product retrieval. While in the real world, the attribute values of a product are usually incomplete and vary over time, which greatly hinders the practical applications. In this paper, we propose a multimodal method to jointly predict product attributes and extract values from textual product descriptions with the help of the product images. We argue that product attributes and values are highly correlated, e.g., it will be easier to extract the values on condition that the product attributes are given. Thus, we jointly model the attribute prediction and value extraction tasks from multiple aspects towards the interactions between attributes and values. Moreover, product images have distinct effects on our tasks for different product attributes and values. Thus, we selectively draw useful visual information from product images to enhance our model. We annotate a multimodal product attribute value dataset that contains 87,194 instances, and the experimental results on this dataset demonstrate that explicitly modeling the relationship between attributes and values facilitates our method to establish the correspondence between them, and selectively utilizing visual product information is necessary for the task. Our code and dataset are available at https://github. com/jd-aig/JAVE .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- EcomGPT: Instruction-Tuning Large Language Models with Chain-of-Task Tasks for E-commerceYangning Li, Shirong Ma, Xiaobin Wang, Shen Huang 等AAAI 2024 · 被引用 85 次
- Inflate and Shrink: Enriching and Reducing Interactions for Fast Text-Image RetrievalHaoliang Liu, Tan Yu, Ping LiEMNLP 2021 · 被引用 14 次
- Jellyfish: Instruction-Tuning Local Large Language Models for Data PreprocessingHaochen Zhang, Yuyang Dong, Chuan Xiao, Masafumi OyamadaEMNLP 2024 · 被引用 11 次
- Product Question Answering in E-Commerce: A SurveyYang Deng, Wenxuan Zhang, Qian Yu, Wai LamACL 2023 · 被引用 9 次
- Multi-Label Zero-Shot Product Attribute-Value ExtractionJiaying Gong, Hoda EldardiryWWW 2024 · 被引用 8 次
它引用的顶会 Paper2
相关 Paper
- Hypergraph-based Zero-shot Multi-modal Product Attribute Value ExtractionJiazhen Hu, Jiaying Gong, Hongda Shen, Hoda EldardiryWWW 2025 · 被引用 4 次
- Open-World Attribute Mining for E-Commerce Products with Multimodal Self-Correction Instruction TuningJiaqi Li, Yanming Li, Xiaoli Shen, Chuanyi Zhang 等ACL 2025 · 被引用 2 次
- JDDC 2.1: A Multimodal Chinese Dialogue Dataset with Joint Tasks of Query Rewriting, Response Generation, Discourse Parsing, and SummarizationNan Zhao, Haoran Li, Youzheng Wu, Xiaodong HeEMNLP 2022 · 被引用 6 次
- Visually Precise QueryRiddhiman Dasgupta, Francis Tom, Sudhir Kumar, Mithun Das Gupta 等ACM MM 2020 · 被引用 1 次
- Price Suggestion for Online Second-hand Items with Texts and ImagesLiang Han, Zhaozheng Yin, Zhurong Xia, Minqian Tang 等ACM MM 2020 · 被引用 8 次
