Domain Disentangled Generative Adversarial Network for Zero-Shot Sketch-Based 3D Shape Retrieval
Rui Xu, Zongyan Han, Le Hui, Jianjun Qian, Jin Xie
摘要
Sketch-based 3D shape retrieval is a challenging task due to the large domain discrepancy between sketches and 3D shapes. Since existing methods are trained and evaluated on the same categories, they cannot effectively recognize the categories that have not been used during training. In this paper, we propose a novel domain disentangled generative adversarial network (DD-GAN) for zero-shot sketch-based 3D retrieval, which can retrieve the unseen categories that are not accessed during training. Specifically, we first generate domain-invariant features and domain-specific features by disentangling the learned features of sketches and 3D shapes, where the domain-invariant features are used to align with the corresponding word embeddings. Then, we develop a generative adversarial network that combines the domain-specific features of the seen categories with the aligned domain-invariant features to synthesize samples, where the synthesized samples of the unseen categories are generated by using the corresponding word embeddings. Finally, we use the synthesized samples of the unseen categories combined with the real samples of the seen categories to train the network for retrieval, so that the unseen categories can be recognized. In order to reduce the domain shift problem, we utilize unlabeled unseen samples to enhance the discrimination ability of the discriminator. With the discriminator distinguishing the generated samples from the unlabeled unseen samples, the generator can generate more realistic unseen samples. Extensive experiments on the SHREC'13 and SHREC'14 datasets show that our method significantly improves the retrieval performance of the unseen categories.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Democratising 2D Sketch to 3D Shape Retrieval Through PivotingPinaki Nath Chowdhury, Ayan Kumar Bhunia, Aneeshan Sain, Subhadeep Koley 等ICCV 2023 · 被引用 10 次
- Fine-grained Prototypical Voting with Heterogeneous Mixup for Semi-supervised 2D-3D Cross-modal RetrievalFan Zhang, Xian-Sheng Hua, Chong Chen, Xiao LuoCVPR 2024 · 被引用 5 次
- DREAM: Decoupled Discriminative Learning with Bigraph-aware Alignment for Semi-supervised 2D-3D Cross-modal RetrievalFan Zhang, Changhu Wang, Zebang Cheng, Xiaojiang Peng 等AAAI 2025 · 被引用 1 次
- What Can Human Sketches Do for Object Detection?Pinaki Nath Chowdhury, Ayan Kumar Bhunia, Aneeshan Sain, Subhadeep Koley 等CVPR 2023
- Doodle Your 3D: from Abstract Freehand Sketches to Precise 3D ShapesHmrishav Bandyopadhyay, Subhadeep Koley, Ayan Das, Ayan Kumar Bhunia 等CVPR 2024
它引用的顶会 Paper3
- Semantic-Aware Knowledge Preservation for Zero-Shot Sketch-Based Image RetrievalQing Liu, Lingxi Xie, Huiyu Wang, Alan L. YuilleICCV 2019 · 被引用 126 次
- Learning the Redundancy-Free Features for Generalized Zero-Shot Object RecognitionZongyan Han, Zhenyong Fu, Jian YangCVPR 2020
- Contrastive Embedding for Generalized Zero-Shot LearningZongyan Han, Zhenyong Fu, Shuo Chen, Jian YangCVPR 2021
相关 Paper
- Learning Cross-Aligned Latent Embeddings for Zero-Shot Cross-Modal RetrievalKaiyi Lin, Xing Xu, Lianli Gao, Zheng Wang 等AAAI 2020 · 被引用 50 次
- Correlated Features Synthesis and Alignment for Zero-shot Cross-modal RetrievalXing Xu, Kaiyi Lin, Huimin Lu, Lianli Gao 等SIGIR 2020 · 被引用 22 次
- Domain-Specific Alignment Network for Multi-Domain Image-Based 3D Object RetrievalYuting Su, Yuqian Li, Dan Song, Zhendong Mao 等ACM MM 2020 · 被引用 3 次
- Multimodal Disentanglement Variational AutoEncoders for Zero-Shot Cross-Modal RetrievalJialin Tian, Kai Wang, Xing Xu, Zuo Cao 等SIGIR 2022 · 被引用 19 次
- GTNet: Generative Transfer Network for Zero-Shot Object DetectionShizhen Zhao, Changxin Gao, Yuanjie Shao, Lerenhan Li 等AAAI 2020 · 被引用 64 次
