Modeling 3D Layout For Group Re-Identification
Quan Zhang, Kaiheng Dang, Jian-Huang Lai, Zhan-Xiang Feng, Xiaohua Xie
Abstract
Group re-identification (GReID) attempts to correctly associate groups with the same members under different cameras. The main challenge is how to resist the membership and layout variations. Existing works attempt to incorporate layout modeling on the basis of appearance features to achieve robust group representations. However, layout ambiguity is introduced because these methods only consider the 2D layout on the imaging plane. In this paper, we overcome the above limitations by 3D layout modeling. Specifically, we propose a novel 3D transformer (3DT) that reconstructs the relative 3D layout relationship among members, then applies sampling and quantification to preset a series of layout tokens along three dimensions, and selects the corresponding tokens as layout features for each member. Furthermore, we build a synthetic GReID dataset, City1M, including 1.84M images, 45K persons and 11.5K groups with 3D annotations to alleviate data shortages and poor annotations. To the best of our knowledge, 3DT is the first work to address GReID with 3D perspective, and the City1M is the currently largest dataset. Several experiments show the superiority of our 3DT and City1M. Our project has been released on https://github.com/ LinlyAC/City1M-dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- MMA: Multi-Modal Adapter for Vision-Language ModelsLingxiao Yang, Ru-Yuan Zhang, Yanchen Wang, Xiaohua XieCVPR 2024 · 46 citations
- View-decoupled Transformer for Person Re-identification under Aerial-ground Camera NetworkQuan Zhang, Lei Wang, Vishal M. Patel, Xiaohua Xie et al.CVPR 2024 · 30 citations
- SEAS: ShapE-Aligned Supervision for Person Re-IdentificationHaidong Zhu, Pranav Budhwant, Zhaoheng Zheng, Ram NevatiaCVPR 2024 · 14 citations
- View-Aware Semantic Alignment for Aerial-Ground Person Re-IdentificationQuan Zhang, Zeqiang Cai, Peiming Zhao, Jingze Wu et al.CVPR 2026 · 1 citation
- Similarity Metric Learning For RGB-Infrared Group Re-IdentificationJianghao Xiong, Jianhuang LaiCVPR 2023
Builds on6
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- TransReID: Transformer-based Object Re-IdentificationShuting He, Hao Luo, Pichao Wang, Fan Wang et al.ICCV 2021 · 1,172 citations
- Surpassing Real-World Source Training Data: Random 3D Characters for Generalizable Person Re-IdentificationYanan Wang, Shengcai Liao, Ling ShaoACM MM 2020 · 90 citations
- AdaBins: Depth Estimation Using Adaptive BinsShariq Farooq Bhat, Ibraheem Alhashim, Peter WonkaCVPR 2021
- Google Landmarks Dataset v2 - A Large-Scale Benchmark for Instance-Level Recognition and RetrievalTobias Weyand, André Araújo, Bingyi Cao, Jack SimCVPR 2020
Related papers
- Uncertainty Modeling with Second-Order Transformer for Group Re-identificationQuan Zhang, Jian-Huang Lai, Zhan-Xiang Feng, Xiaohua XieAAAI 2022 · 22 citations
- Can Person-Level Attributes Improve Group Re-Identification?Kamakshya Prasad Nayak, Kamalakar Vijay Thakare, Ashesh Xalxo, Lalit Lohani et al.ACM MM 2025
- AG-VPReID: A Challenging Large-Scale Benchmark for Aerial-Ground Video-based Person Re-IdentificationHuy Nguyen, Kien Nguyen, Akila Pemasiri, Feng Liu et al.CVPR 2025
- Zillow Indoor Dataset: Annotated Floor Plans With 360deg Panoramas and 3D Room LayoutsSteve Cruz, Will Hutchcroft, Yuguang Li, Naji Khosravan et al.CVPR 2021
- BV-Person: A Large-scale Dataset for Bird-view Person Re-identificationCheng Yan, Guansong Pang, Lei Wang, Jile Jiao et al.ICCV 2021 · 26 citations
