Mathematical Justification of Hard Negative Mining via Isometric Approximation Theorem
Albert Xu, Jhih-Yi Hsieh, Bhaskar Vundurthy, Nithya Kemp, Eliana Cohen, Lu Li, Howie Choset
Abstract
In deep metric learning, the Triplet Loss has emerged as a popular method to learn many computer vision and natural language processing tasks such as facial recognition, object detection, and visual-semantic embeddings. One issue that plagues the Triplet Loss is network collapse, an undesirable phenomenon where the network projects the embeddings of all data onto a single point. Researchers predominately solve this problem by using triplet mining strategies. While hard negative mining is the most effective of these strategies, existing formulations lack strong theoretical justification for their empirical success. In this paper, we utilize the mathematical theory of isometric approximation to show an equivalence between the Triplet Loss sampled by hard negative mining and an optimization problem that minimizes a Hausdorff-like distance between the neural network and its ideal counterpart function. This provides the theoretical justifications for hard negative mining's empirical efficacy. In addition, our novel application of the isometric approximation theorem provides the groundwork for future forms of hard negative mining that avoid network collapse. Our theory can also be extended to analyze other Euclidean space-based metric learning methods like Ladder Loss or Contrastive Learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 58f4ae15-2baf-49e2-8359-5db7d42d6517Cited by top-tier papers1
Ask how each one uses itBuilds on5
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi et al.NeurIPS 2020 · 2,611 citations
- CSI: Novelty Detection via Contrastive Learning on Distributionally Shifted InstancesJihoon Tack, Sangwoo Mo, Jongheon Jeong, Jinwoo ShinNeurIPS 2020 · 755 citations
- Ladder Loss for Coherent Visual-Semantic EmbeddingMo Zhou, Zhenxing Niu, Le Wang, Zhanning Gao et al.AAAI 2020 · 46 citations
- Rethinking preventing class-collapsing in metric learning with margin-based lossesElad Levi, Tete Xiao, Xiaolong Wang, Trevor DarrellICCV 2021 · 15 citations
Related papers
- SoftTriple Loss: Deep Metric Learning Without Triplet SamplingQi Qian, Lei Shang, Baigui Sun, Juhua Hu et al.ICCV 2019 · 419 citations
- LoOp: Looking for Optimal Hard Negative Embeddings for Deep Metric LearningBhavya Vasudeva, Puneesh Deora, Saumik Bhattacharya, Umapada Pal et al.ICCV 2021 · 16 citations
- Deep Metric Learning With Tuplet Margin LossBaosheng Yu, Dacheng TaoICCV 2019 · 104 citations
- The Dilemma of TriHard Loss and an Element-Weighted TriHard Loss for Person Re-IdentificationYihao Lv, Youzhi Gu, Xinggao LiuNeurIPS 2020 · 11 citations
- RankMI: A Mutual Information Maximizing Ranking LossMete Kemertas, Leila Pishdad, Konstantinos G. Derpanis, Afsaneh FazlyCVPR 2020
