Generalized Sum Pooling for Metric Learning
Yeti Ziya Gürbüz, Ozan Sener, A. Aydin Alatan
Abstract
A common architectural choice for deep metric learning is a convolutional neural network followed by global average pooling (GAP). Albeit simple, GAP is a highly effective way to aggregate information. One possible explanation for the effectiveness of GAP is considering each feature vector as representing a different semantic entity and GAP as a convex combination of them. Following this perspective, we generalize GAP and propose a learnable generalized sum pooling method (GSP). GSP improves GAP with two distinct abilities: i) the ability to choose a subset of semantic entities, effectively learning to ignore nuisance information, and ii) learning the weights corresponding to the importance of each entity. Formally, we propose an entropy-smoothed optimal transport problem and show that it is a strict generalization of GAP, i.e., a specific realization of the problem gives back GAP. We show that this optimization problem enjoys analytical gradients enabling us to use it as a direct learnable replacement for GAP. We further propose a zero-shot loss to ease the learning of GSP. We show the effectiveness of our method with extensive evaluations on 4 popular metric learning benchmarks. Code is available at: GSP-DML Framework
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d60bc45c-5275-422e-b74f-bffe91a41b6fBuilds on31
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- EfficientNetV2: Smaller Models and Faster TrainingMingxing Tan, Quoc V. LeICML 2021 · 4,239 citations
- Attribute Prototype Network for Zero-Shot LearningWenjia Xu, Yongqin Xian, Jiuniu Wang, Bernt Schiele et al.NeurIPS 2020 · 392 citations
- Revisiting Training Strategies and Generalization Performance in Deep Metric LearningKarsten Roth, Timo Milbich, Samarth Sinha, Prateek Gupta et al.ICML 2020 · 187 citations
- Differentiable Top-k with Optimal TransportYujia Xie, Hanjun Dai, Minshuo Chen, Bo Dai et al.NeurIPS 2020 · 124 citations
Related papers
- Deep Disentangled Metric LearningJinhee Park, Jisoo Park, Dagyeong Na, Junseok KwonAAAI 2025 · 3 citations
- Deep Compositional Metric LearningWenzhao Zheng, Chengkun Wang, Jiwen Lu, Jie ZhouCVPR 2021
- Deep Metric Learning via Adaptive Learnable AssessmentWenzhao Zheng, Jiwen Lu, Jie ZhouCVPR 2020
- Cross-Image-Attention for Conditional Embeddings in Deep Metric LearningDmytro Kotovenko, Pingchuan Ma, Timo Milbich, Björn OmmerCVPR 2023
- Towards Interpretable Deep Metric Learning with Structural MatchingWenliang Zhao, Yongming Rao, Ziyi Wang, Jiwen Lu et al.ICCV 2021 · 52 citations
