Causality Compensated Attention for Contextual Biased Visual Recognition
Ruyang Liu, Jingjia Huang, Thomas H. Li, Ge Li
Abstract
Visual attention does not always capture the essential object representation desired for robust predictions. Attention modules tend to underline not only the target object but also the common co-occurring context that the module thinks helpful in the training. The problem is rooted in the confounding effect of the context leading to incorrect causalities between objects and predictions, which is further exacerbated by visual attention. In this paper, to learn causal object features robust for contextual bias, we propose a novel attention module named Interventional Dual Attention (IDA) for visual recognition. Specifically, IDA adopts two attention layers with multiple sampling intervention, which compensates the attention against the confounder context. Note that our method is model-agnostic and thus can be implemented on various backbones. Extensive experiments show our model obtains significant improvements in classification and detection with lower computation. In particular, we achieve the state-of-the-art results in multi-label classification on MS-COCO and PASCAL-VOC.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 2a379ce3-ad5c-4bd9-bc58-baefa29d577fCited by top-tier papers5
- Counterfactual Reasoning for Multi-Label Image Classification via Patching-Based TrainingMing-Kun Xie, Jiahao Xiao, Pei Peng, Gang Niu et al.ICML 2024 · 11 citations
- In Pursuit of Causal Label Correlations for Multi-label Image RecognitionZhao-Min Chen, Xin Jin, Yisu Ge, Sixian ChanNeurIPS 2024 · 9 citations
- Causal-Entity Reflected Egocentric Traffic Accident Video SynthesisLei-Lei Li, Jianwu Fang, Junbin Xiao, Shanmin Pang et al.ICCV 2025 · 4 citations
- Prototype-based Causal Intervention for Multi-Label Image ClassificationYanmin Li, Zhilong Mao, Mao Wang, Lihua Liu et al.CVPR 2026
- Correlative and Discriminative Label Grouping for Multi-Label Visual Prompt TuningLei-Lei Ma, Shuo Xu, Ming-Kun Xie, Lei Wang et al.CVPR 2025
Related papers
- Contextual Debiasing for Visual Recognition with Causal MechanismsRuyang Liu, Hao Liu, Ge Li, Haodi Hou et al.CVPR 2022 · 42 citations
- Causal Attention for Unbiased Visual RecognitionTan Wang, Chang Zhou, Qianru Sun, Hanwang ZhangICCV 2021 · 162 citations
- Improving Weakly Supervised Object Localization via Causal InterventionFeifei Shao, Yawei Luo, Li Zhang, Lu Ye et al.ACM MM 2021 · 24 citations
- Global Meets Local: Effective Multi-Label Image Classification via Category-Aware Weak SupervisionJiawei Zhan, Jun Liu, Wei Tang, Guannan Jiang et al.ACM MM 2022 · 6 citations
- Deep Contextual Attention for Human-Object Interaction DetectionTiancai Wang, Rao Muhammad Anwer, Muhammad Haris Khan, Fahad Shahbaz Khan et al.ICCV 2019 · 130 citations
