Cross-Modal Center Loss for 3D Cross-Modal Retrieval
Longlong Jing, Elahe Vahdani, Jiaxing Tan, Yingli Tian
摘要
Cross-modal retrieval aims to learn discriminative and modal-invariant features for data from different modalities. Unlike the existing methods which usually learn from the features extracted by offline networks, in this paper, we propose an approach to jointly train the components of crossmodal retrieval framework with metadata, and enable the network to find optimal features. The proposed end-to-end framework is updated with three loss functions: 1) a novel cross-modal center loss to eliminate cross-modal discrepancy, 2) cross-entropy loss to maximize inter-class variations, and 3) mean-square-error loss to reduce modality variations. In particular, our proposed cross-modal center loss minimizes the distances of features from objects belonging to the same class across all modalities. Extensive experiments have been conducted on the retrieval tasks across multi-modalities including 2D image, 3D point cloud and mesh data. The proposed framework significantly outperforms the state-of-the-art methods for both cross-modal and in-domain retrieval for 3D objects on the ModelNet10 and ModelNet40 datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Contrast with Reconstruct: Contrastive 3D Representation Learning Guided by Generative PretrainingZekun Qi, Runpei Dong, Guofan Fan, Zheng Ge 等ICML 2023 · 被引用 209 次
- ReconBoost: Boosting Can Achieve Modality ReconcilementCong Hua, Qianqian Xu, Shilong Bao, Zhiyong Yang 等ICML 2024 · 被引用 52 次
- Single Image 3D Shape Retrieval via Cross-Modal Instance and Category Contrastive LearningMing-Xian Lin, Jie Yang, He Wang, Yu-Kun Lai 等ICCV 2021 · 被引用 34 次
- Autoencoders as Cross-Modal Teachers: Can Pretrained 2D Image Transformers Help 3D Representation Learning?Runpei Dong, Zekun Qi, Linfeng Zhang, Junbo Zhang 等ICLR 2023 · 被引用 21 次
- Semi-supervised Knowledge Transfer Across Multi-omic Single-cell DataFan Zhang, Tianyu Liu, Zihao Chen, Xiaojiang Peng 等NeurIPS 2024 · 被引用 7 次
它引用的顶会 Paper4
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna 等NeurIPS 2020 · 被引用 7,049 次
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui 等ICCV 2019 · 被引用 3,193 次
- DeepGCNs: Can GCNs Go As Deep As CNNs?Guohao Li, Matthias Müller, Ali K. Thabet, Bernard GhanemICCV 2019 · 被引用 1,586 次
- View N-Gram Network for 3D Object RetrievalXinwei He, Tengteng Huang, Song Bai, Xiang BaiICCV 2019 · 被引用 65 次
相关 Paper
- C3CMR: Cross-Modality Cross-Instance Contrastive Learning for Cross-Media RetrievalJunsheng Wang, Tiantian Gong, Zhixiong Zeng, Changchang Sun 等ACM MM 2022 · 被引用 12 次
- Multi-graph Convolutional Network for Unsupervised 3D Shape RetrievalWeizhi Nie, Yue Zhao, An-An Liu, Zan Gao 等ACM MM 2020 · 被引用 10 次
- MSO: Multi-Feature Space Joint Optimization Network for RGB-Infrared Person Re-IdentificationYajun Gao, Tengfei Liang, Yi Jin, Xiaoyan Gu 等ACM MM 2021 · 被引用 75 次
- MCCN: Multimodal Coordinated Clustering Network for Large-Scale Cross-modal RetrievalZhixiong Zeng, Ying Sun, Wenji MaoACM MM 2021 · 被引用 20 次
- Cross-Modality Person Re-Identification via Modality Confusion and Center AggregationXin Hao, Sanyuan Zhao, Mang Ye, Jianbing ShenICCV 2021 · 被引用 191 次
