Modeling 3D Layout For Group Re-Identification
Quan Zhang, Kaiheng Dang, Jian-Huang Lai, Zhan-Xiang Feng, Xiaohua Xie
摘要
Group re-identification (GReID) attempts to correctly associate groups with the same members under different cameras. The main challenge is how to resist the membership and layout variations. Existing works attempt to incorporate layout modeling on the basis of appearance features to achieve robust group representations. However, layout ambiguity is introduced because these methods only consider the 2D layout on the imaging plane. In this paper, we overcome the above limitations by 3D layout modeling. Specifically, we propose a novel 3D transformer (3DT) that reconstructs the relative 3D layout relationship among members, then applies sampling and quantification to preset a series of layout tokens along three dimensions, and selects the corresponding tokens as layout features for each member. Furthermore, we build a synthetic GReID dataset, City1M, including 1.84M images, 45K persons and 11.5K groups with 3D annotations to alleviate data shortages and poor annotations. To the best of our knowledge, 3DT is the first work to address GReID with 3D perspective, and the City1M is the currently largest dataset. Several experiments show the superiority of our 3DT and City1M. Our project has been released on https://github.com/ LinlyAC/City1M-dataset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- MMA: Multi-Modal Adapter for Vision-Language ModelsLingxiao Yang, Ru-Yuan Zhang, Yanchen Wang, Xiaohua XieCVPR 2024 · 被引用 46 次
- View-decoupled Transformer for Person Re-identification under Aerial-ground Camera NetworkQuan Zhang, Lei Wang, Vishal M. Patel, Xiaohua Xie 等CVPR 2024 · 被引用 30 次
- SEAS: ShapE-Aligned Supervision for Person Re-IdentificationHaidong Zhu, Pranav Budhwant, Zhaoheng Zheng, Ram NevatiaCVPR 2024 · 被引用 14 次
- View-Aware Semantic Alignment for Aerial-Ground Person Re-IdentificationQuan Zhang, Zeqiang Cai, Peiming Zhao, Jingze Wu 等CVPR 2026 · 被引用 1 次
- Similarity Metric Learning For RGB-Infrared Group Re-IdentificationJianghao Xiong, Jianhuang LaiCVPR 2023
它引用的顶会 Paper6
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- TransReID: Transformer-based Object Re-IdentificationShuting He, Hao Luo, Pichao Wang, Fan Wang 等ICCV 2021 · 被引用 1,172 次
- Surpassing Real-World Source Training Data: Random 3D Characters for Generalizable Person Re-IdentificationYanan Wang, Shengcai Liao, Ling ShaoACM MM 2020 · 被引用 90 次
- AdaBins: Depth Estimation Using Adaptive BinsShariq Farooq Bhat, Ibraheem Alhashim, Peter WonkaCVPR 2021
- Google Landmarks Dataset v2 - A Large-Scale Benchmark for Instance-Level Recognition and RetrievalTobias Weyand, André Araújo, Bingyi Cao, Jack SimCVPR 2020
相关 Paper
- Uncertainty Modeling with Second-Order Transformer for Group Re-identificationQuan Zhang, Jian-Huang Lai, Zhan-Xiang Feng, Xiaohua XieAAAI 2022 · 被引用 22 次
- Can Person-Level Attributes Improve Group Re-Identification?Kamakshya Prasad Nayak, Kamalakar Vijay Thakare, Ashesh Xalxo, Lalit Lohani 等ACM MM 2025
- AG-VPReID: A Challenging Large-Scale Benchmark for Aerial-Ground Video-based Person Re-IdentificationHuy Nguyen, Kien Nguyen, Akila Pemasiri, Feng Liu 等CVPR 2025
- Zillow Indoor Dataset: Annotated Floor Plans With 360deg Panoramas and 3D Room LayoutsSteve Cruz, Will Hutchcroft, Yuguang Li, Naji Khosravan 等CVPR 2021
- BV-Person: A Large-scale Dataset for Bird-view Person Re-identificationCheng Yan, Guansong Pang, Lei Wang, Jile Jiao 等ICCV 2021 · 被引用 26 次
