Deconfounding Physical Dynamics with Global Causal Relation and Confounder Transmission for Counterfactual Prediction
Zongzhao Li, Xiangyu Zhu, Zhen Lei, Zhaoxiang Zhang
Abstract
Discovering the underneath causal relations is the fundamental ability for reasoning about the surrounding environment and predicting the future states in the physical world. Counterfactual prediction from visual input, which requires simulating future states based on unrealized situations in the past, is a vital component in causal relation tasks. In this paper, we work on the confounders that have effect on the physical dynamics, including masses, friction coefficients, etc., to bridge relations between the intervened variable and the affected variable whose future state may be altered. We propose a neural network framework combining Global Causal Relation Attention (GCRA) and Confounder Transmission Structure (CTS). The GCRA looks for the latent causal relations between different variables and estimates the confounders by capturing both spatial and temporal information. The CTS integrates and transmits the learnt confounders in a residual way, so that the estimated confounders can be encoded into the network as a constraint for object positions when performing counterfactual prediction. Without any access to ground truth information about confounders, our model outperforms the state-of-the-art method on various benchmarks by fully utilizing the constraints of confounders. Extensive experiments demonstrate that our model can generalize to unseen environments and maintain good performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0629fac8-9b77-43dc-b7cb-edaebeff6588Cited by top-tier papers1
Ask how each one uses itBuilds on8
- CLEVRER: Collision Events for Video Representation and ReasoningKexin Yi, Chuang Gan, Yunzhu Li, Pushmeet Kohli et al.ICLR 2020 · 584 citations
- Learning Compositional Koopman Operators for Model-Based ControlYunzhu Li, Hao He, Jiajun Wu, Dina Katabi et al.ICLR 2020 · 135 citations
- Causal Discovery in Physical Systems from VideosYunzhu Li, Antonio Torralba, Anima Anandkumar, Dieter Fox et al.NeurIPS 2020 · 133 citations
- CoPhy: Counterfactual Learning of Physical DynamicsFabien Baradel, Natalia Neverova, Julien Mille, Greg Mori et al.ICLR 2020 · 105 citations
- Attention over Learned Object Embeddings Enables Complex Visual ReasoningDavid Ding, Felix Hill, Adam Santoro, Malcolm Reynolds et al.NeurIPS 2021 · 87 citations
Related papers
- Filtered-CoPhy: Unsupervised Learning of Counterfactual Physics in Pixel SpaceSteeven Janny, Fabien Baradel, Natalia Neverova, Madiha Nadri et al.ICLR 2022 · 18 citations
- Disentangled Counterfactual Learning for Physical Audiovisual Commonsense ReasoningChangsheng Lv, Shuai Zhang, Yapeng Tian, Mengshi Qi et al.NeurIPS 2023 · 26 citations
- Grounding Physical Concepts of Objects and Events Through Dynamic Visual ReasoningZhenfang Chen, Jiayuan Mao, Jiajun Wu, Kwan-Yee Kenneth Wong et al.ICLR 2021 · 13 citations
- Learning Causal Representation for Training Cross-Domain Pose Estimator via Generative InterventionsXiheng Zhang, Yongkang Wong, Xiaofei Wu, Juwei Lu et al.ICCV 2021 · 36 citations
- Causal Discovery and Inference through Next-Token PredictionEivinas Butkus, Nikolaus KriegeskorteNeurIPS 2025 · 3 citations
