On the Variability of Concept Activation Vectors
Julia Wenkmann, Damien Garreau
摘要
One of the most pressing challenges in artificial intelligence is to make models more transparent to their users. Recently, explainable artificial intelligence has come up with numerous methods to tackle this challenge. A promising avenue is to use concept-based explanations, that is, high-level concepts instead of plain feature importance scores. Among this class of methods, Concept Activation Vectors (CAVs, Kim et al., 2018) stand out as one of the main protagonists. One interesting aspect of CAVs is that their computation requires sampling random examples from the train set. Therefore, the actual vectors obtained may vary depending on the randomness of this sampling. In this paper, we propose a fine-grained theoretical analysis of CAV construction in order to quantify their variability. Our results, confirmed by experiments on several real-life datasets of four different modalities, point to an universal result: the variance of CAVs declines roughly as , where is the number of random examples. Based on this, we give practical recommendations for a resource-efficient application of the method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper9
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- Invertible Concept-based Explanations for CNN Models with Non-negative Concept Activation VectorsRuihan Zhang, Prashan Madumal, Tim Miller, Krista A. Ehinger 等AAAI 2021 · 被引用 140 次
- CoCoX: Generating Conceptual and Counterfactual Explanations via Fault-LinesArjun R. Akula, Shuai Wang, Song-Chun ZhuAAAI 2020 · 被引用 102 次
- Concept Activation Regions: A Generalized Framework For Concept-Based ExplanationsJonathan Crabbé, Mihaela van der SchaarNeurIPS 2022 · 被引用 88 次
- What does LIME really see in images?Damien Garreau, Dina MardaouiICML 2021 · 被引用 49 次
相关 Paper
- FastCAV: Efficient Computation of Concept Activation Vectors for Explaining Deep Neural NetworksLaines Schmalwasser, Niklas Penzel, Joachim Denzler, Julia NieblingICML 2025
- Concept Gradient: Concept-based Interpretation Without Linear AssumptionAndrew Bai, Chih-Kuan Yeh, Neil Y. C. Lin, Pradeep Kumar Ravikumar 等ICLR 2023 · 被引用 5 次
- Navigating Neural Space: Revisiting Concept Activation Vectors to Overcome Directional DivergenceFrederik Pahde, Maximilian Dreyer, Moritz Weckbecker, Leander Weber 等ICLR 2025
- Concept Distillation: Leveraging Human-Centered Explanations for Model ImprovementAvani Gupta, Saurabh Saini, P. J. NarayananNeurIPS 2023 · 被引用 18 次
- LG-CAV: Train Any Concept Activation Vector with Language GuidanceQihan Huang, Jie Song, Mengqi Xue, Haofei Zhang 等NeurIPS 2024 · 被引用 12 次
