Learning Causal Representation for Training Cross-Domain Pose Estimator via Generative Interventions
Xiheng Zhang, Yongkang Wong, Xiaofei Wu, Juwei Lu, Mohan S. Kankanhalli, Xiangdong Li, Weidong Geng
Abstract
3D pose estimation has attracted increasing attention with the availability of high-quality benchmark datasets. However, prior works show that deep learning models tend to learn spurious correlations, which fail to generalize beyond the specific dataset they are trained on. In this work, we take a step towards training robust models for cross-domain pose estimation task, which brings together ideas from causal representation learning and generative adversarial networks. Specifically, this paper introduces a novel framework for causal representation learning which explicitly exploits the causal structure of the task. We consider changing domain as interventions on images under the data-generation process and steer the generative model to produce counterfactual features. This help the model learn transferable and causal relations across different domains. Our framework is able to learn with various types of unlabeled datasets. We demonstrate the efficacy of our proposed method on both human and hand pose estimation task. The experiment results show the proposed approach achieves state-of-the-art performance on most datasets for both domain adaptation and domain generalization settings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d7d6572b-7f79-44e4-b378-3997e7f12ffdCited by top-tier papers7
- Lagrange Motion Analysis and View Embeddings for Improved Gait RecognitionTianrui Chai, Annan Li, Shaoxiong Zhang, Zilong Li et al.CVPR 2022 · 84 citations
- Uncertainty-Aware Adaptation for Self-Supervised 3D Human Pose EstimationJogendra Nath Kundu, Siddharth Seth, Pradyumna YM, Varun Jampani et al.CVPR 2022 · 41 citations
- Learning Event-Relevant Factors for Video Anomaly DetectionChe Sun, Chenrui Shi, Yunde Jia, Yuwei WuAAAI 2023 · 16 citations
- CPL: Counterfactual Prompt Learning for Vision and Language ModelsXuehai He, Diji Yang, Weixi Feng, Tsu-Jui Fu et al.EMNLP 2022 · 13 citations
- Motif-Consistent Counterfactuals with Adversarial Refinement for Graph-level Anomaly DetectionChunjing Xiao, Shikang Pang, Wenxin Tai, Yanlong Huang et al.KDD 2024 · 5 citations
Builds on8
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- FreiHAND: A Dataset for Markerless Capture of Hand Pose and Shape From Single RGB ImagesChristian Zimmermann, Duygu Ceylan, Jimei Yang, Bryan C. Russell et al.ICCV 2019 · 493 citations
- Representation Learning via Invariant Causal MechanismsJovana Mitrovic, Brian McWilliams, Jacob C. Walker, Lars Holger Buesing et al.ICLR 2021 · 281 citations
- Counterfactual Generative NetworksAxel Sauer, Andreas GeigerICLR 2021 · 145 citations
- Inference Stage Optimization for Cross-scenario 3D Human Pose EstimationJianfeng Zhang, Xuecheng Nie, Jiashi FengNeurIPS 2020 · 53 citations
Related papers
- Generative Interventions for Causal LearningChengzhi Mao, Augustine Cha, Amogh Gupta, Hao Wang et al.CVPR 2021
- Causal Transportability for Visual RecognitionChengzhi Mao, Kevin Xia, James Wang, Hao Wang et al.CVPR 2022 · 27 citations
- Invariant and Transportable Representations for Anti-Causal Domain ShiftsYibo Jiang, Victor VeitchNeurIPS 2022 · 50 citations
- Counterfactual Maximum Likelihood Estimation for Training Deep NetworksXinyi Wang, Wenhu Chen, Michael Saxon, William Yang WangNeurIPS 2021 · 9 citations
- Learning Robust Intervention Representations with Delta EmbeddingsPanagiotis Alimisis, Christos DiouICLR 2026
