Counterfactual Vision and Language Learning
Ehsan Abbasnejad, Damien Teney, Amin Parvaneh, Javen Shi, Anton van den Hengel
摘要
The ongoing success of visual question answering methods has been somewhat surprising given that, at its most general, the problem requires understanding the entire variety of both visual and language stimuli. It is particularly remarkable that this success has been achieved on the basis of comparatively small datasets, given the scale of the problem. One explanation is that this has been accomplished partly by exploiting bias in the datasets rather than developing deeper multi-modal reasoning. This fundamentally limits the generalization of the method, and thus its practical applicability. We propose a method that addresses this problem by introducing counterfactuals in the training. In doing so we leverage structural causal models for counterfactual evaluation to formulate alternatives, for instance, questions that could be asked of the same image set. We show that simulating plausible alternative training data through this process results in better generalization.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper40
- On the Value of Out-of-Distribution Testing: An Example of Goodhart's LawDamien Teney, Ehsan Abbasnejad, Kushal Kafle, Robik Shrestha 等NeurIPS 2020 · 被引用 163 次
- Counterfactual Data-Augmented Sequential RecommendationZhenlei Wang, Jingsen Zhang, Hongteng Xu, Xu Chen 等SIGIR 2021 · 被引用 131 次
- Family as a Third Space for AI Literacies: How do children and parents learn about AI together?Stefania Druga, Fee Lia Christoph, Amy J. KoCHI 2022 · 被引用 118 次
- Active Learning by Feature MixingAmin Parvaneh, Ehsan Abbasnejad, Damien Teney, Reza Haffari 等CVPR 2022 · 被引用 113 次
- Cross-View Geo-Localization via Learning Disentangled Geometric Layout CorrespondenceXiaohan Zhang, Xingyu Li, Waqas Sultani, Yi Zhou 等AAAI 2023 · 被引用 111 次
相关 Paper
- Counterfactual VQA: A Cause-Effect Look at Language BiasYulei Niu, Kaihua Tang, Hanwang Zhang, Zhiwu Lu 等CVPR 2021
- Mitigating Language Bias of LMMs in Social Intelligence Understanding with Virtual Counterfactual CalibrationPeng Chen, Xiao-Yu Guo, Yuan-Fang Li, Xiaowang Zhang 等EMNLP 2024 · 被引用 1 次
- CF-VLM: CounterFactual Vision-Language Fine-tuningJusheng Zhang, Kaitong Cai, Yijia Fan, Jian Wang 等NeurIPS 2025 · 被引用 71 次
- What If the TV was off? Examining Counterfactual Reasoning Abilities of Multi-modal Language ModelsLetian Zhang, Xiaotong Zhai, Zhongkai Zhao, Yongshuo Zong 等CVPR 2024 · 被引用 9 次
- On the General Value of Evidence, and Bilingual Scene-Text Visual Question AnsweringXinyu Wang, Yuliang Liu, Chunhua Shen, Chun Chet Ng 等CVPR 2020
