Generative multitask learning mitigates target-causing confounding
Taro Makino, Krzysztof J. Geras, Kyunghyun Cho
Abstract
We propose generative multitask learning (GMTL), a simple and scalable approach to causal representation learning for multitask learning. Our approach makes a minor change to the conventional multitask inference objective, and improves robustness to target shift. Since GMTL only modifies the inference objective, it can be used with existing multitask learning methods without requiring additional training. The improvement in robustness comes from mitigating unobserved confounders that cause the targets, but not the input. We refer to them as target-causing confounders. These confounders induce spurious dependencies between the input and targets. This poses a problem for conventional multitask learning, due to its assumption that the targets are conditionally independent given the input. GMTL mitigates target-causing confounding at inference time, by removing the influence of the joint target distribution, and predicting all targets jointly. This removes the spurious dependencies between the input and targets, where the degree of removal is adjustable via a single hyperparameter. This flexibility is useful for managing the trade-off between in- and out-of-distribution generalization. Our results on the Attributes of People and Taskonomy datasets reflect an improved robustness to target shift across four multitask learning methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Beyond Invariance: Test-Time Label-Shift Adaptation for Addressing "Spurious" CorrelationsQingyao Sun, Kevin P. Murphy, Sayna Ebrahimi, Alexander D'AmourNeurIPS 2023 · 10 citations
- Identifying and Mitigating Spurious Correlation in Multi-Task LearningJunyi Chai, Shenyu Lu, Xiaoqian WangCVPR 2025
Builds on6
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 1,416 citations
- Which Tasks Should Be Learned Together in Multi-task Learning?Trevor Standley, Amir Zamir, Dawn Chen, Leonidas J. Guibas et al.ICML 2020 · 651 citations
- Robust fine-tuning of zero-shot modelsMitchell Wortsman, Gabriel Ilharco, Jong Wook Kim, Mike Li et al.CVPR 2022 · 364 citations
- Muppet: Massive Multi-task Representations with Pre-FinetuningArmen Aghajanyan, Anchit Gupta, Akshat Shrivastava, Xilun Chen et al.EMNLP 2021 · 176 citations
- Identifying Statistical Bias in Dataset ReplicationLogan Engstrom, Andrew Ilyas, Shibani Santurkar, Dimitris Tsipras et al.ICML 2020 · 55 citations
Related papers
- Improving Multi-Task Generalization via Regularizing Spurious CorrelationZiniu Hu, Zhe Zhao, Xinyang Yi, Tiansheng Yao et al.NeurIPS 2022 · 46 citations
- C-Disentanglement: Discovering Causally-Independent Generative Factors under an Inductive Bias of ConfounderXiaoyu Liu, Jiaxin Yuan, Bang An, Yuancheng Xu et al.NeurIPS 2023 · 13 citations
- Generative Interventions for Causal LearningChengzhi Mao, Augustine Cha, Amogh Gupta, Hao Wang et al.CVPR 2021
- Learning Causal Representation for Training Cross-Domain Pose Estimator via Generative InterventionsXiheng Zhang, Yongkang Wong, Xiaofei Wu, Juwei Lu et al.ICCV 2021 · 36 citations
- Causal Transportability for Visual RecognitionChengzhi Mao, Kevin Xia, James Wang, Hao Wang et al.CVPR 2022 · 27 citations
