Targeted Activation Penalties Help CNNs Ignore Spurious Signals
Dekai Zhang, Matt Williams, Francesca Toni
Abstract
Neural networks (NNs) can learn to rely on spurious signals in the training data, leading to poor generalisation. Recent methods tackle this problem by training NNs with additional ground-truth annotations of such signals. These methods may, however, let spurious signals re-emerge in deep convolutional NNs (CNNs). We propose Targeted Activation Penalty (TAP), a new method tackling the same problem by penalising activations to control the re-emergence of spurious signals in deep CNNs, while also lowering training times and memory usage. In addition, ground-truth annotations can be expensive to obtain. We show that TAP still works well with annotations generated by pre-trained models as effective substitutes of ground-truth annotations. We demonstrate the power of TAP against two state-of-the-art baselines on the MNIST benchmark and on two clinical image datasets, using four different CNN architectures.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d0171233-4ab7-4bfb-acfb-e836d68eb901Builds on3
- Interpretations are Useful: Penalizing Explanations to Align Neural Networks with Prior KnowledgeLaura Rieger, Chandan Singh, W. James Murdoch, Bin YuICML 2020 · 249 citations
- Post hoc Explanations may be Ineffective for Detecting Unknown Spurious CorrelationJulius Adebayo, Michael Muelly, Harold Abelson, Been KimICLR 2022 · 102 citations
- Right for Better Reasons: Training Differentiable Models by Constraining their Influence FunctionsXiaoting Shao, Arseny Skryagin, Wolfgang Stammer, Patrick Schramowski et al.AAAI 2021 · 45 citations
Related papers
- Overcoming Simplicity Bias in Deep Networks using a Feature SieveRishabh Tiwari, Pradeep ShenoyICML 2023 · 32 citations
- Skin Deep Unlearning: Artefact and Instrument Debiasing in the Context of Melanoma ClassificationPeter J. Bevan, Amir Atapour-AbarghoueiICML 2022 · 24 citations
- MotionTTT: 2D Test-Time-Training Motion Estimation for 3D Motion Corrected MRITobit Klug, Kun Wang, Stefan Ruschke, Reinhard HeckelNeurIPS 2024 · 4 citations
- Let Samples Speak: Mitigating Spurious Correlation by Exploiting the Clusterness of SamplesWeiwei Li, Junzhuo Liu, Yuanyuan Ren, Yuchen Zheng et al.CVPR 2025
- JPEG-ACT: Accelerating Deep Learning via Transform-based Lossy CompressionR. David Evans, Lufei Liu, Tor M. AamodtISCA 2020 · 47 citations
