Semantic Consistency Guided Instance Feature Alignment for 2D Image-Based 3D Shape Retrieval
Heyu Zhou, Weizhi Nie, Dan Song, Nian Hu, Xuanya Li, An-An Liu
Abstract
2D image-based 3D shape retrieval (2D-to-3D) investigates the problem of matching the relevant 3D shapes from gallery dataset when given a query image. Recently, adversarial training and environmental style transfer learning have been successful applied to this task and achieved state-of-the-art performance. However, there still exist two problems. First, previous works only concentrate on the connection between the label and representation, where the unique visual characteristics of each instance are paid less attention. Second, the confused features or the transformed images can only cheat the discriminator but can not guarantee the semantic consistency. In another words, features of 2D desk may be mapped nearby the features of 3D chair. In this paper, we propose a novel semantic consistency guided instance feature alignment network (SC-IFA) to address these limitations. SC-IFA mainly consists of two parts, instance visual feature extraction and cross-domain instance feature adaptation. For the first module, unlike previous methods, which merely employ 2D CNN to extract the feature, we additionally maximize the mutual information between the input and feature to enhance the capability of feature representation for each instance. For the second module, we first introduce the margin disparity discrepancy model to mix up the cross-domain features in an adversarial training way. Then, we design two feature translators to transform the feature from one domain to another domain, and impose the translation loss and correlation loss on the transformed features to preserve the semantic consistency. Extensive experimental results on two benchmarks, MI3DOR and MI3DOR-2, verify SC-IFA is superior to the state-of-the-art methods.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 64d68c36-49a8-4d85-acd2-71b86aa158c9Cited by top-tier papers1
Ask how each one uses itRelated papers
- Cross-Domain 3D Model Retrieval Based On Contrastive Learning And Label PropagationDan Song, Yue Yang, Weizhi Nie, Xuanya Li et al.ACM MM 2022 · 6 citations
- Unsupervised 2D Image-Based 3D Model Retrieval via Decision Boundary Alignment and Graph Semantic PropagationNian Hu, Yibo Zhao, Xinhui Li, Chen Li et al.SIGIR 2026
- Domain-Specific Alignment Network for Multi-Domain Image-Based 3D Object RetrievalYuting Su, Yuqian Li, Dan Song, Zhendong Mao et al.ACM MM 2020 · 3 citations
- Instance-Guided Scene Adaptation for Unsupervised Person SearchLinfeng Qi, Huibing Wang, Jinjia Peng, Xianping Fu et al.AAAI 2026
- Single Image 3D Shape Retrieval via Cross-Modal Instance and Category Contrastive LearningMing-Xian Lin, Jie Yang, He Wang, Yu-Kun Lai et al.ICCV 2021 · 34 citations
