AG-VPReID: A Challenging Large-Scale Benchmark for Aerial-Ground Video-based Person Re-Identification
Huy Nguyen, Kien Nguyen, Akila Pemasiri, Feng Liu, Sridha Sridharan, Clinton Fookes
摘要
We introduce AG-VPReID, a new large-scale dataset for aerial-ground video-based person re-identification (ReID) that comprises 6,632 subjects, 32,321 tracklets and over 9.6 million frames captured by drones (altitudes ranging from 15-120m), CCTV, and wearable cameras. This dataset offers a real-world benchmark for evaluating the robustness to significant viewpoint changes, scale variations, and resolution differences in cross-platform aerial-ground settings. In addition, to address these challenges, we propose AG-VPReID-Net, an end-to-end framework composed of three complementary streams: (1) an Adapted Temporal-Spatial Stream addressing motion pattern inconsistencies and facilitating temporal feature learning, (2) a Normalized Appearance Stream leveraging physics-informed techniques to tackle resolution and appearance changes, and (3) a Multi-Scale Attention Stream handling scale variations across drone altitudes. We integrate visual-semantic cues from all streams to form a robust, viewpoint-invariant whole-body representation. Extensive experiments demonstrate that AG-VPReID-Net outperforms state-of-the-art approaches on both our new dataset and existing video-based ReID benchmarks, showcasing its effectiveness and generalizability. Nevertheless, the performance gap observed on AG-VPReID across all methods underscores the dataset's challenging nature. The dataset, code and trained models are available at AG-VPReID-Net.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Try Harder: Hard Sample Generation and Learning for Cloth-Changing Person Re-IDHankun Liu, Yujian Zhao, Guanglin NiuACM MM 2025 · 被引用 1 次
- Cross-modal Fuzzy Alignment Network for Text-Aerial Person Retrieval and A Large-scale BenchmarkYifei Deng, Chenglong Li, Yuyang Zhang, Guyue Hu 等CVPR 2026
- Semantic-Driven Visual Progressive Refinement for Aerial-Ground Person ReID: A Challenging Large-Scale BenchmarkAihua Zheng, Hao Xie, Xixi Wan, Zi Wang 等AAAI 2026
它引用的顶会 Paper20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- CLIP-ReID: Exploiting Vision-Language Model for Image Re-identification without Concrete Text LabelsSiyuan Li, Li Sun, Qingli LiAAAI 2023 · 被引用 355 次
- Global-Local Temporal Representations for Video Person Re-IdentificationJianing Li, Shiliang Zhang, Jingdong Wang, Wen Gao 等ICCV 2019 · 被引用 241 次
- Clothes-Changing Person Re-identification with RGB Modality OnlyXinqian Gu, Hong Chang, Bingpeng Ma, Shutao Bai 等CVPR 2022 · 被引用 226 次
- Pyramid Spatial-Temporal Aggregation for Video-based Person Re-IdentificationYingquan Wang, Pingping Zhang, Shang Gao, Xia Geng 等ICCV 2021 · 被引用 118 次
相关 Paper
- View-Aware Semantic Alignment for Aerial-Ground Person Re-IdentificationQuan Zhang, Zeqiang Cai, Peiming Zhao, Jingze Wu 等CVPR 2026 · 被引用 1 次
- SeCap: Self-Calibrating and Adaptive Prompts for Cross-view Person Re-Identification in Aerial-Ground NetworksShining Wang, Yunlong Wang, Ruiqi Wu, Bingliang Jiao 等CVPR 2025
- GSAlign: Geometric and Semantic Alignment Network for Aerial-Ground Person Re-IdentificationQiao Li, Jie Li, Yukang Zhang, Lei Tan 等NeurIPS 2025 · 被引用 5 次
- View-decoupled Transformer for Person Re-identification under Aerial-ground Camera NetworkQuan Zhang, Lei Wang, Vishal M. Patel, Xiaohua Xie 等CVPR 2024 · 被引用 30 次
- BV-Person: A Large-scale Dataset for Bird-view Person Re-identificationCheng Yan, Guansong Pang, Lei Wang, Jile Jiao 等ICCV 2021 · 被引用 26 次
