Multi-perspective Coherent Reasoning for Helpfulness Prediction of Multimodal Reviews
Junhao Liu, Zhen Hai, Min Yang, Lidong Bing
摘要
As more and more product reviews are posted in both text and images, Multimodal Review Analysis (MRA) becomes an attractive research topic. Among the existing review analysis tasks, helpfulness prediction on review text has become predominant due to its importance for e-commerce platforms and online shops, i.e. helping customers quickly acquire useful product information. This paper proposes a new task Multimodal Review Helpfulness Prediction (MRHP) aiming to analyze the review helpfulness from text and visual modalities. Meanwhile, a novel Multi-perspective Coherent Reasoning method (MCR) is proposed to solve the MRHP task, which conducts joint reasoning over texts and images from both the product and the review, and aggregates the signals to predict the review helpfulness. Concretely, we first propose a productreview coherent reasoning module to measure the intra-and inter-modal coherence between the target product and the review. In addition, we also devise an intra-review coherent reasoning module to identify the coherence between the text content and images of the review, which is a piece of strong evidence for review helpfulness prediction. To evaluate the effectiveness of MCR, we present two newly collected multimodal review datasets as benchmark evaluation resources for the MRHP task. Experimental results show that our MCR method can lead to a performance increase of up to 8.5% as compared to the best performing text-only model. The source code and datasets can be obtained from https:// github.com/jhliu17/MCR.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
相关 Paper
- Adaptive Contrastive Learning on Multimodal Transformer for Review Helpfulness PredictionThong Nguyen, Xiaobao Wu, Anh Tuan Luu, Zhen Hai 等EMNLP 2022 · 被引用 8 次
- A Multi-Modal Context Reasoning Approach for Conditional Inference on Joint Textual and Visual CluesYunxin Li, Baotian Hu, Xinyu Chen, Yuxin Ding 等ACL 2023 · 被引用 10 次
- Premise-based Multimodal Reasoning: Conditional Inference on Joint Textual and Visual CluesQingxiu Dong, Ziwei Qin, Heming Xia, Tian Feng 等ACL 2022
- Multimodal Joint Attribute Prediction and Value Extraction for E-commerce ProductTiangang Zhu, Yue Wang, Haoran Li, Youzheng Wu 等EMNLP 2020 · 被引用 46 次
- A Survey of Multimodal Mathematical Reasoning: From Perception, Alignment to ReasoningTianyu Yang, Sihong Wu, Yilun Zhao, Zhenwen Liang 等ACL 2026
