Addressing Leakage in Concept Bottleneck Models
Marton Havasi, Sonali Parbhoo, Finale Doshi-Velez
摘要
Concept bottleneck models (CBMs) enhance the interpretability of their predictions by first predicting high-level concepts given features, and subsequently predicting outcomes on the basis of these concepts. Recently, it was demonstrated that training the label predictor directly on the probabilities produced by the concept predictor as opposed to the ground-truth concepts, improves label predictions. However, this results in corruptions in the concept predictions that impact the concept accuracy as well as our ability to intervene on the concepts -a key proposed benefit of CBMs. In this work, we investigate and address two issues with CBMs that cause this disparity in performance: having an insufficient concept set and using inexpressive concept predictor. With our modifications, CBMs become competitive in terms of predictive performance, with models that otherwise leak unintended information in the concept probabilities, while having dramatically increased concept accuracy and intervention accuracy.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper47
- Probabilistic Concept Bottleneck ModelsEunji Kim, Dahuin Jung, Sangha Park, Siwon Kim 等ICML 2023 · 被引用 108 次
- A Closer Look at the Intervention Procedure of Concept Bottleneck ModelsSungbin Shin, Yohan Jo, Sungsoo Ahn, Namhoon LeeICML 2023 · 被引用 59 次
- Learning to Receive Help: Intervention-Aware Concept Embedding ModelsMateo Espinosa Zarlenga, Katie Collins, Krishnamurthy Dvijotham, Adrian Weller 等NeurIPS 2023 · 被引用 56 次
- Stochastic Concept Bottleneck ModelsMoritz Vandenhirtz, Sonia Laguna, Ricards Marcinkevics, Julia E. VogtNeurIPS 2024 · 被引用 56 次
- Auxiliary Losses for Learning Generalizable Concept-based ModelsIvaxi Sheth, Samira Ebrahimi KahouNeurIPS 2023 · 被引用 52 次
它引用的顶会 Paper2
相关 Paper
- Learning to Intervene on Concept BottlenecksDavid Steinmann, Wolfgang Stammer, Felix Friedrich, Kristian KerstingICML 2024 · 被引用 32 次
- Concepts' Information Bottleneck ModelsKarim Galliamov, Syed Muhammad Ahsan Raza Kazmi, Adil Khan, Adín Ramírez RiveraICLR 2026 · 被引用 2 次
- An Analysis of Concept Bottleneck Models: Measuring, Understanding, and Mitigating the Impact of Noisy AnnotationsSeonghwan Park, Jueun Mun, Donghyun Oh, Namhoon LeeNeurIPS 2025 · 被引用 10 次
- There Was Never a Bottleneck in Concept Bottleneck ModelsAntonio Almudévar, José Miguel Hernández-Lobato, Alfonso OrtegaICLR 2026 · 被引用 9 次
- Interactive Concept Bottleneck ModelsKushal Chauhan, Rishabh Tiwari, Jan Freyberg, Pradeep Shenoy 等AAAI 2023 · 被引用 91 次
