StyleMeUp: Towards Style-Agnostic Sketch-Based Image Retrieval
Aneeshan Sain, Ayan Kumar Bhunia, Yongxin Yang, Tao Xiang, Yi-Zhe Song
Abstract
Sketch-based image retrieval (SBIR) is a cross-modal matching problem which is typically solved by learning a joint embedding space where the semantic content shared between photo and sketch modalities are preserved. However, a fundamental challenge in SBIR has been largely ignored so far, that is, sketches are drawn by humans and considerable style variations exist amongst different users. An effective SBIR model needs to explicitly account for this style diversity, crucially, to generalise to unseen user styles. To this end, a novel style-agnostic SBIR model is proposed. Different from existing models, a cross-modal variational autoencoder (VAE) is employed to explicitly disentangle each sketch into a semantic content part shared with the corresponding photo, and a style part unique to the sketcher. Importantly, to make our model dynamically adaptable to any unseen user styles, we propose to metatrain our cross-modal VAE by adding two style-adaptive components: a set of feature transformation layers to its encoder and a regulariser to the disentangled semantic content latent code. With this meta-learning framework, our model can not only disentangle the cross-modal shared semantic content for SBIR, but can adapt the disentanglement to any unseen user style as well, making the SBIR model truly style-agnostic. Extensive experiments show that our style-agnostic model yields state-of-the-art performance for both category-level and instance-level SBIR.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3ad5aa2f-62c6-4214-89d7-b49945edb0c0Cited by top-tier papers8
- GenQuery: Supporting Expressive Visual Search with Generative ModelsKihoon Son, DaEun Choi, Tae Soo Kim, Young-Ho Kim et al.CHI 2024 · 48 citations
- Correspondence-Free Domain Alignment for Unsupervised Cross-Domain Image RetrievalXu Wang, Dezhong Peng, Ming Yan, Peng HuAAAI 2023 · 35 citations
- Unsupervised Cross-Domain Image Retrieval via Prototypical Optimal TransportBin Li, Ye Shi, Qian Yu, Jingya WangAAAI 2024 · 16 citations
- DiDA: Disambiguated Domain Alignment for Cross-Domain Retrieval with Partial LabelsHaoran Liu, Ying Ma, Ming Yan, Yingke Chen et al.AAAI 2024 · 13 citations
- Differentiable Auxiliary Learning for Sketch Re-IdentificationXingyu Liu, Xu Cheng, Haoyu Chen, Hao Yu et al.AAAI 2024 · 11 citations
Builds on8
- Cross-Domain Few-Shot Classification via Learned Feature-Wise TransformationHung-Yu Tseng, Hsin-Ying Lee, Jia-Bin Huang, Ming-Hsuan YangICLR 2020 · 467 citations
- Content and Style Disentanglement for Artistic Style TransferDmytro Kotovenko, Artsiom Sanakoyeu, Sabine Lang, Björn OmmerICCV 2019 · 187 citations
- Goal-Driven Sequential Data AbstractionUmar Riaz Muhammad, Yongxin Yang, Timothy M. Hospedales, Tao Xiang et al.ICCV 2019 · 25 citations
- More Photos Are All You Need: Semi-Supervised Learning for Fine-Grained Sketch Based Image RetrievalAyan Kumar Bhunia, Pinaki Nath Chowdhury, Aneeshan Sain, Yongxin Yang et al.CVPR 2021
- Sketch Less for More: On-the-Fly Fine-Grained Sketch-Based Image RetrievalAyan Kumar Bhunia, Yongxin Yang, Timothy M. Hospedales, Tao Xiang et al.CVPR 2020
Related papers
- Scene-Level Sketch-Based Image Retrieval with Minimal Pairwise SupervisionCe Ge, Jingyu Wang, Qi Qi, Haifeng Sun et al.AAAI 2023 · 6 citations
- Sketch3T: Test-Time Training for Zero-Shot SBIRAneeshan Sain, Ayan Kumar Bhunia, Vaishnav Potlapalli, Pinaki Nath Chowdhury et al.CVPR 2022 · 55 citations
- Multimodal Disentanglement Variational AutoEncoders for Zero-Shot Cross-Modal RetrievalJialin Tian, Kai Wang, Xing Xu, Zuo Cao et al.SIGIR 2022 · 19 citations
- How to Handle Sketch-Abstraction in Sketch-Based Image Retrieval?Subhadeep Koley, Ayan Kumar Bhunia, Aneeshan Sain, Pinaki Nath Chowdhury et al.CVPR 2024
- Semi-transductive Learning for Generalized Zero-Shot Sketch-Based Image RetrievalCe Ge, Jingyu Wang, Qi Qi, Haifeng Sun et al.AAAI 2023 · 10 citations
