Hierarchical classification at multiple operating points
Jack Valmadre
Abstract
Many classification problems consider classes that form a hierarchy. Classifiers that are aware of this hierarchy may be able to make confident predictions at a coarse level despite being uncertain at the fine-grained level. While it is generally possible to vary the granularity of predictions using a threshold at inference time, most contemporary work considers only leaf-node prediction, and almost no prior work has compared methods at multiple operating points. We present an efficient algorithm to produce operating characteristic curves for any method that assigns a score to every class in the hierarchy. Applying this technique to evaluate existing methods reveals that top-down classifiers are dominated by a naïve flat softmax classifier across the entire operating range. We further propose two novel loss functions and show that a soft variant of the structured hinge loss is able to significantly outperform the flat baseline. Finally, we investigate the poor accuracy of top-down classifiers and demonstrate that they perform relatively well on unseen classes. Code is available online at https://github.com/jvlmdr/hiercls . 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e2eb940f-7f7f-4fda-9aad-795f6a32b685Cited by top-tier papers7
- Hierarchical Selective ClassificationShani Goren, Ido Galil, Ran El-YanivNeurIPS 2024 · 16 citations
- LCA-on-the-Line: Benchmarking Out of Distribution Generalization with Class TaxonomiesJia Shi, Gautam Rajendrakumar Gare, Jinjin Tian, Siqi Chai et al.ICML 2024 · 13 citations
- Test-Time Amendment with a Coarse Classifier for Fine-Grained ClassificationKanishk Jain, Shyamgopal Karthik, Vineet GandhiNeurIPS 2023 · 9 citations
- To Each Metric Its Decoding: Post-Hoc Optimal Decision Rules of Probabilistic Hierarchical ClassifiersRoman Plaud, Alexandre Perez-Lebel, Matthieu Labeau, Antoine Saillenfest et al.ICML 2025
- Flattening the Parent Bias: Hierarchical Semantic Segmentation in the Poincaré BallSimon Weber, Baris Zöngür, Nikita Araslanov, Daniel CremersCVPR 2024
Builds on6
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Long-tail learning via logit adjustmentAditya Krishna Menon, Sadeep Jayasumana, Ankit Singh Rawat, Himanshu Jain et al.ICLR 2021 · 937 citations
- No Cost Likelihood Manipulation at Test Time for Making Better Mistakes in Deep NetworksShyamgopal Karthik, Ameya Prabhu, Puneet K. Dokania, Vineet GandhiICLR 2021 · 26 citations
- Making Better Mistakes: Leveraging Class Hierarchies With Deep NetworksLuca Bertinetto, Romain Müller, Konstantinos Tertikas, Sina Samangooei et al.CVPR 2020
Related papers
- Hierarchical Entity Typing via Multi-level Learning to RankTongfei Chen, Yunmo Chen, Benjamin Van DurmeACL 2020 · 51 citations
- Stochastic smoothing of the top-K calibrated hinge loss for deep imbalanced classificationCamille Garcin, Maximilien Servajean, Alexis Joly, Joseph SalmonICML 2022 · 14 citations
- Training Uncertainty-Aware Classifiers with Conformalized Deep LearningBat-Sheva Einbinder, Yaniv Romano, Matteo Sesia, Yanfei ZhouNeurIPS 2022 · 84 citations
- Label Relation Graphs Enhanced Hierarchical Residual Network for Hierarchical Multi-Granularity ClassificationJingzhou Chen, Peng Wang, Jian Liu, Yuntao QianCVPR 2022 · 56 citations
- Hier-COS: Making Deep Features Hierarchy-aware via Composition of Orthogonal SubspacesDepanshu Sani, Saket AnandCVPR 2026
