Learning Attribute-driven Disentangled Representations for Interactive Fashion Retrieval
Yuxin Hou, Eleonora Vig, Michael Donoser, Loris Bazzani
摘要
Interactive retrieval for online fashion shopping provides the ability to change image retrieval results according to the user feedback. One common problem in interactive retrieval is that a specific user interaction (e.g., changing the color of a T-shirt) causes other aspects to change inadvertently (e.g., the retrieved item has a sleeve type different than the query). This is a consequence of existing methods learning visual representations that are semantically entangled in the embedding space, which limits the controllability of the retrieved results. We propose to leverage on the semantics of visual attributes to train convolutional networks that learn attribute-specific subspaces for each attribute to obtain disentangled representations. Thus operations, such as swapping out a particular attribute value for another, impact the attribute at hand and leave others untouched. We show that our model can be tailored to deal with different retrieval tasks while maintaining its disentanglement property. We obtain state-of-the-art performance on three interactive fashion retrieval tasks: attribute manipulation retrieval, conditional similarity retrieval, and outfit complementary item retrieval. Code and models are publicly available1.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Dynamic Weighted Combiner for Mixed-Modal Image RetrievalFuxiang Huang, Lei Zhang, Xiaowei Fu, Suqi SongAAAI 2024 · 被引用 28 次
- MUST: An Effective and Scalable Framework for Multimodal Search of Target ModalityMengzhao Wang, Xiangyu Ke, Xiaoliang Xu, Lu Chen 等ICDE 2024 · 被引用 16 次
- FashionNTM: Multi-turn Fashion Image Retrieval via Cascaded MemoryAnwesan Pal, Sahil Wadhwa, Ayush Jaiswal, Xu Zhang 等ICCV 2023 · 被引用 11 次
- Multi-modal Extreme ClassificationAnshul Mittal, Kunal Dahiya, Shreya Malani, Janani Ramaswamy 等CVPR 2022 · 被引用 10 次
- Identifying Ambiguous Similarity Conditions via Semantic MatchingHan-Jia Ye, Yi Shi, De-Chuan ZhanCVPR 2022 · 被引用 5 次
它引用的顶会 Paper13
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 被引用 2,416 次
- Swapping Autoencoder for Deep Image ManipulationTaesung Park, Jun-Yan Zhu, Oliver Wang, Jingwan Lu 等NeurIPS 2020 · 被引用 376 次
- CamNet: Coarse-to-Fine Retrieval for Camera Re-LocalizationMingyu Ding, Zhe Wang, Jiankai Sun, Jianping Shi 等ICCV 2019 · 被引用 163 次
- Monocular Neural Image Based Rendering With Continuous View ControlJie Song, Xu Chen, Otmar HilligesICCV 2019 · 被引用 85 次
- Attribute Manipulation Generative Adversarial Networks for Fashion ImagesKenan E. Ak, Ashraf A. Kassim, Joo-Hwee Lim, Jo Yew ThamICCV 2019 · 被引用 85 次
相关 Paper
- DiSCo: Disentangled Attribute Manipulation Retrieval via Semantic Reconstruction and Consistency RegularizationMin Tan, Guanhao Liu, Huijing Zhan, Yuyu Yin 等ACM MM 2025
- Conditional Cross Attention Network for Multi-Space Embedding without Entanglement in Only a SINGLE NetworkChull Hwan Song, Taebaek Hwang, Jooyoung Yoon, Shunghyun Choi 等ICCV 2023 · 被引用 2 次
- Controllable Gradient Item RetrievalHaonan Wang, Chang Zhou, Carl Yang, Hongxia Yang 等WWW 2021 · 被引用 11 次
- Learning Attribute and Class-Specific Representation Duet for Fine-Grained Fashion AnalysisYang Jiao, Yan Gao, Jingjing Meng, Jin Shang 等CVPR 2023
- Generative Attribute Manipulation Scheme for Flexible Fashion SearchXin Yang, Xuemeng Song, Xianjing Han, Haokun Wen 等SIGIR 2020 · 被引用 32 次
