Very Important Person Localization in Unconstrained Conditions: A New Benchmark
Xiao Wang, Zheng Wang, Toshihiko Yamasaki, Wenjun Zeng
Abstract
This paper presents a new high-quality dataset for Very Important Person Localization (VIPLoc), named Unconstrained-7k. Generally, current datasets: 1) are limited in scale; 2) built under simple and constrained conditions, where the number of disturbing non-VIPs is not large, the scene is relatively simple, and the face of VIP is always in frontal view and salient. To tackle these problems, the proposed Unconstrained-7k dataset is featured in two aspects. First, it contains over 7,000 annotated images, making it the largest VIPLoc dataset under unconstrained conditions to date. Second, our dataset is collected freely on the Internet, including multiple scenes, where images are in unconstrained conditions. VIPs in the new dataset are in different settings, e.g., large view variation, varying sizes, occluded, and complex scenes. Meanwhile, each image has more persons (> 20), making the dataset more challenging.
As a minor contribution, motivated by the observation that VIPs are highly related to not only neighbors but also iconic objects, this paper proposes a Joint Social Relation and Individual Interaction Graph Neural Networks (JSRII-GNN) for VIPLoc. Experiments show that the JSRII-GNN yields competitive accuracy on NCAA (National Collegiate Athletic Association), MS (Multi-scene), and Unconstrained-7k datasets. https://github.com/xiaowang1516/VIPLoc.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 72eee921-51cc-4d98-9f44-c8ed9f08ceedCited by top-tier papers1
Ask how each one uses itBuilds on9
- Measuring and Relieving the Over-Smoothing Problem for Graph Neural Networks from the Topological ViewDeli Chen, Yankai Lin, Wei Li, Peng Li et al.AAAI 2020 · 1,353 citations
- Schema-Guided Multi-Domain Dialogue State Tracking with Graph Attention Neural NetworksLu Chen, Boer Lv, Chi Wang, Su Zhu et al.AAAI 2020 · 143 citations
- Beyond the Parts: Learning Multi-view Cross-part Correlation for Vehicle Re-identificationXinchen Liu, Wu Liu, Jinkai Zheng, Chenggang Yan et al.ACM MM 2020 · 97 citations
- AATEAM: Achieving the Ad Hoc Teamwork by Employing the Attention MechanismShuo Chen, Ewa Andrejczuk, Zhiguang Cao, Jie ZhangAAAI 2020 · 54 citations
- Image Enhanced Event Detection in News ArticlesMeihan Tong, Shuai Wang, Yixin Cao, Bin Xu et al.AAAI 2020 · 43 citations
Related papers
- Towards Causality Inference for Very Important Person LocalizationXiao Wang, Zheng Wang, Wu Liu, Xin Xu et al.ACM MM 2022 · 2 citations
- Crowd3D: Towards Hundreds of People Reconstruction from a Single ImageHao Wen, Jing Huang, Huili Cui, Haozhe Lin et al.CVPR 2023
- Detecting and Grounding Important Characters in Visual StoriesDanyang Liu, Frank KellerAAAI 2023 · 11 citations
- Video Individual Counting for Moving DronesYaowu Fan, Jia Wan, Tao Han, Antoni B. Chan et al.ICCV 2025 · 1 citation
- Understanding Human Gaze Communication by Spatio-Temporal Graph ReasoningLifeng Fan, Wenguan Wang, Song-Chun Zhu, Xinyu Tang et al.ICCV 2019 · 124 citations
