Towards Causality Inference for Very Important Person Localization
Xiao Wang, Zheng Wang, Wu Liu, Xin Xu, Qijun Zhao, Shin'ichi Satoh
Abstract
Very Important Person Localization (VIPLoc) aims at detecting certain individuals in a given image, who are more attractive than others in the image. Existing uncontrolled VIPLoc benchmark assumes that the image has one single VIP, which is not suitable for actual application scenarios when multiple VIPs or no VIPs appear in the image. In this paper, we re-built a complex uncontrolled conditions (CUC) dataset to make the VIPLoc closer to the actual situation, containing no, single, and multiple VIPs. Existing methods use the hand-designed and deep learning strategies to extract the features of persons and analyze the differences between VIPs and other persons from the perspective of statistics. They are not explainable as to why the VIP located this output for that input. Thus, there exist the severe performance degradation when we use these models in real-world VIPLoc. Specifically, we establish a causal inference framework that unpacks the causes of previous methods and derives a new principled solution for VIPLoc. It treats the scene as confounding factor, allowing the ever-elusive confounding effects to be eliminated and the essential determinants to be uncovered. Through extensive experiments, our method outperforms the state-of-the-art methods on public VIPLoc datasets and the re-built CUC dataset.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 5534a658-686a-4a18-b865-be6878b861eaCited by top-tier papers1
Ask how each one uses itRelated papers
- Very Important Person Localization in Unconstrained Conditions: A New BenchmarkXiao Wang, Zheng Wang, Toshihiko Yamasaki, Wenjun ZengAAAI 2021 · 11 citations
- Crowd3D: Towards Hundreds of People Reconstruction from a Single ImageHao Wen, Jing Huang, Huili Cui, Haozhe Lin et al.CVPR 2023
- CrossHOI-Bench: A Unified Benchmark for HOI Evaluation across Vision-Language Models and HOI-Specific MethodsQinqian Lei, Bo Wang, Robby T. TanCVPR 2026 · 6 citations
- Multimodal Causal Reasoning for UAV Object DetectionNianxin Li, Mao Ye, Lihua Zhou, Shuaifeng Li et al.NeurIPS 2025 · 1 citation
- A Causal Inference Look at Unsupervised Video Anomaly DetectionXiangru Lin, Yuyang Chen, Guanbin Li, Yizhou YuAAAI 2022 · 44 citations
