Switchable Representation Learning Framework with Self-Compatibility
Shengsen Wu, Yan Bai, Yihang Lou, Xiongkun Linghu, Jianzhong He, Ling-Yu Duan
Abstract
Real-world visual search systems involve deployments on multiple platforms with different computing and storage resources. Deploying a unified model that suits the minimal-constrain platforms leads to limited accuracy. It is expected to deploy models with different capacities adapting to the resource constraints, which requires features extracted by these models to be aligned in the metric space. The method to achieve feature alignments is called "compatible learning". Existing research mainly focuses on the one-to-one compatible paradigm, which is limited in learning compatibility among multiple models. We propose a Switchable representation learning Framework with Self-Compatibility (SFSC). SFSC generates a series of compatible sub-models with different capacities through one training process. The optimization of sub-models faces gradients conflict, and we mitigate this problem from the perspective of the magnitude and direction. We adjust the priorities of sub-models dynamically through uncertainty estimation to co-optimize sub-models properly. Besides, the gradients with conflicting directions are projected to avoid mutual interference. SFSC achieves state-of-the-art performance on the evaluated datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bd2dca8e-c549-41f8-8bb0-0e1f5786c917Cited by top-tier papers1
Ask how each one uses itBuilds on17
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine et al.NeurIPS 2020 · 2,261 citations
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
- LeViT: a Vision Transformer in ConvNet's Clothing for Faster InferenceBenjamin Graham, Alaaeldin El-Nouby, Hugo Touvron, Pierre Stock et al.ICCV 2021 · 1,009 citations
- Evidential Deep Learning for Open Set Action RecognitionWentao Bao, Qi Yu, Yu KongICCV 2021 · 204 citations
- Glance and Focus: a Dynamic Approach to Reducing Spatial Redundancy in Image ClassificationYulin Wang, Kangchen Lv, Rui Huang, Shiji Song et al.NeurIPS 2020 · 179 citations
Related papers
- Towards Backward-Compatible Representation LearningYantao Shen, Yuanjun Xiong, Wei Xia, Stefano SoattoCVPR 2020
- Learning Compatible EmbeddingsQiang Meng, Chixiang Zhang, Xiaoqiang Xu, Feng ZhouICCV 2021 · 43 citations
- Compatibility-Aware Heterogeneous Visual SearchRahul Duggal, Hao Zhou, Shuo Yang, Yuanjun Xiong et al.CVPR 2021
- Stationary Representations: Optimally Approximating Compatibility and Implications for Improved Model ReplacementsNiccolò Biondi, Federico Pernici, Simone Ricci, Alberto Del BimboCVPR 2024
- A General Rank Preserving Framework for Asymmetric Image RetrievalHui Wu, Min Wang, Wengang Zhou, Houqiang LiICLR 2023
