Multimodal Compatibility Modeling via Exploring the Consistent and Complementary Correlations
Weili Guan, Haokun Wen, Xuemeng Song, Chung-Hsing Yeh, Xiaojun Chang, Liqiang Nie
摘要
Existing methods towards outfit compatibility modeling seldom explicitly consider multimodal correlations. In this work, we explore the consistent and complementary correlations for better compatibility modeling. This is, however, non-trivial due to the following challenges: 1) how to separate and model these two kinds of correlations; 2) how to leverage the derived complementary cues to strengthen the text and vision-oriented representations of the given item; and 3) how to reinforce the compatibility modeling with text and vision-oriented representations. To address these challenges, we present a comprehensive multimodal outfit compatibility modeling scheme. It first nonlinearly projects each modality into separable consistent and complementary spaces via multi-layer perceptron, and then models the consistent and complementary correlations between two modalities by parallel and orthogonal regularization. Thereafter, we strengthen the visual and textual representation of items with complementary information, and further induct both the text-oriented and vision- oriented outfit compatibility modeling. We ultimately employ the mutual learning strategy to reinforce the final performance of compatibility modeling. Extensive experiments demonstrate the superiority of our scheme.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Cosmo: contrastive fusion learning with small data for multimodal human activity recognitionXiaomin Ouyang, Xian Shuai, Jiayu Zhou, Ivy Wang Shi 等MobiCom 2022 · 被引用 94 次
- Target-Guided Composed Image RetrievalHaokun Wen, Xian Zhang, Xuemeng Song, Yinwei Wei 等ACM MM 2023 · 被引用 53 次
- Personalized Fashion Compatibility Modeling via Metapath-guided Heterogeneous Graph LearningWeili Guan, Fangkai Jiao, Xuemeng Song, Haokun Wen 等SIGIR 2022 · 被引用 51 次
- MART: Masked Affective RepresenTation Learning via Masked Temporal Distribution DistillationZhicheng Zhang, Pancheng Zhao, Eunil Park, Jufeng YangCVPR 2024 · 被引用 11 次
- RaCMC: Residual-Aware Compensation Network with Multi-Granularity Constraints for Fake News DetectionXinquan Yu, Ziqi Sheng, Wei Lu, Xiangyang Luo 等AAAI 2025 · 被引用 9 次
它引用的顶会 Paper5
- Mutual Mean-Teaching: Pseudo Label Refinery for Unsupervised Domain Adaptation on Person Re-identificationYixiao Ge, Dapeng Chen, Hongsheng LiICLR 2020 · 被引用 651 次
- Hierarchical Fashion Graph Network for Personalized Outfit RecommendationXingchen Li, Xiang Wang, Xiangnan He, Long Chen 等SIGIR 2020 · 被引用 124 次
- Learning Similarity Conditions Without Explicit SupervisionReuben Tan, Mariya I. Vasileva, Kate Saenko, Bryan A. PlummerICCV 2019 · 被引用 90 次
- Comprehensive Linguistic-Visual Composition Network for Image RetrievalHaokun Wen, Xuemeng Song, Xin Yang, Yibing Zhan 等SIGIR 2021 · 被引用 72 次
- Fashion Compatibility Modeling through a Multi-modal Try-on-guided SchemeXue Dong, Jianlong Wu, Xuemeng Song, Hongjun Dai 等SIGIR 2020 · 被引用 29 次
相关 Paper
- Complementary Factorization towards Outfit Compatibility ModelingTianyu Su, Xuemeng Song, Na Zheng, Weili Guan 等ACM MM 2021 · 被引用 18 次
- Explainable Multi-Modality Alignment for Transferable RecommendationShenghao Yang, Weizhi Ma, Zhiqiang Guo, Min Zhang 等WWW 2025 · 被引用 3 次
- Text2Outfit: Controllable Outfit Generation With Multimodal Language ModelsYuanhao Zhai, Yen-Liang Lin, Minxu Peng, Larry S. Davis 等ICCV 2025 · 被引用 1 次
- Collocation and Try-on Network: Whether an Outfit is CompatibleNa Zheng, Xuemeng Song, Qingying Niu, Xue Dong 等ACM MM 2021 · 被引用 23 次
- Fashion Outfit Complementary Item RetrievalYen-Liang Lin, Son Dinh Tran, Larry S. DavisCVPR 2020
