Auxiliary Losses for Learning Generalizable Concept-based Models
Ivaxi Sheth, Samira Ebrahimi Kahou
Abstract
The increasing use of neural networks in various applications has lead to increasing apprehensions, underscoring the necessity to understand their operations beyond mere final predictions. As a solution to enhance model transparency, Concept Bottleneck Models (CBMs) have gained popularity since their introduction. CBMs essentially limit the latent space of a model to human-understandable high-level concepts. While beneficial, CBMs have been reported to often learn irrelevant concept representations that consecutively damage model performance. To overcome the performance trade-off, we propose cooperative-Concept Bottleneck Model (coop-CBM). The concept representation of our model is particularly meaningful when fine-grained concept labels are absent. Furthermore, we introduce the concept orthogonal loss (COL) to encourage the separation between the concept representations and to reduce the intra-concept distance. This paper presents extensive experiments on real-world datasets for image classification tasks, namely CUB, AwA2, CelebA and TIL. We also study the performance of coop-CBM models under various distributional shift settings. We show that our proposed method achieves higher accuracy in all distributional shift settings even compared to the black-box models with the highest concept accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2acbe257-556a-4e40-854c-080abd7834a2Cited by top-tier papers16
- Concept Bottleneck Generative ModelsAya Abdelsalam Ismail, Julius Adebayo, Héctor Corrada Bravo, Stephen Ra et al.ICLR 2024 · 43 citations
- Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate ExpertsAndrea Pugnana, Riccardo Massidda, Francesco Giannini, Pietro Barbiero et al.NeurIPS 2025 · 11 citations
- An Analysis of Concept Bottleneck Models: Measuring, Understanding, and Mitigating the Impact of Noisy AnnotationsSeonghwan Park, Jueun Mun, Donghyun Oh, Namhoon LeeNeurIPS 2025 · 10 citations
- Towards Multi-dimensional Explanation Alignment for Medical ClassificationLijie Hu, Songning Lai, Wenshuo Chen, Hongru Xiao et al.NeurIPS 2024 · 8 citations
- Sample-efficient Learning of Concepts with Theoretical Guarantees: from Data to Concepts without InterventionsHidde Fokkema, Tim van Erven, Sara MagliacaneNeurIPS 2025 · 7 citations
Builds on15
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Concept Bottleneck ModelsPang Wei Koh, Thao Nguyen, Yew Siang Tang, Stephen Mussmann et al.ICML 2020 · 1,233 citations
- Addressing Leakage in Concept Bottleneck ModelsMarton Havasi, Sonali Parbhoo, Finale Doshi-VelezNeurIPS 2022 · 163 citations
Related papers
- Discovering Fine-Grained Visual-Concept Relations by Disentangled Optimal Transport Concept Bottleneck ModelsYan Xie, Zequn Zeng, Hao Zhang, Yucheng Ding et al.CVPR 2025
- Intervening in Black Box: Concept Bottleneck Model for Enhancing Human Neural Network Mutual UnderstandingNuoye Xiong, Anqi Dong, Ning Wang, Cong Hua et al.ICCV 2025 · 1 citation
- There Was Never a Bottleneck in Concept Bottleneck ModelsAntonio Almudévar, José Miguel Hernández-Lobato, Alfonso OrtegaICLR 2026 · 9 citations
- Label-free Concept Bottleneck ModelsTuomas P. Oikarinen, Subhro Das, Lam M. Nguyen, Tsui-Wei WengICLR 2023 · 17 citations
- Addressing Concept Mislabeling in Concept Bottleneck Models Through Preference OptimizationEmiliano Penaloza, Tianyue H. Zhang, Laurent Charlin, Mateo Espinosa ZarlengaICML 2025
