UAV-Human: A Large Benchmark for Human Behavior Understanding With Unmanned Aerial Vehicles
Tianjiao Li, Jun Liu, Wei Zhang, Yun Ni, Wenqian Wang, Zhiheng Li
摘要
Human behavior understanding with unmanned aerial vehicles (UAVs) is of great significance for a wide range of applications, which simultaneously brings an urgent demand of large, challenging, and comprehensive benchmarks for the development and evaluation of UAV-based models. However, existing benchmarks have limitations in terms of the amount of captured data, types of data modalities, categories of provided tasks, and diversities of subjects and environments. Here we propose a new benchmark -UAV-Human -for human behavior understanding with UAVs, which contains 67,428 multi-modal video sequences and 119 subjects for action recognition, 22,476 frames for pose estimation, 41,290 frames and 1,144 identities for person re-identification, and 22,263 frames for attribute recognition. Our dataset was collected by a flying UAV in multiple urban and rural districts in both daytime and nighttime over three months, hence covering extensive diversities w.r.t subjects, backgrounds, illuminations, weathers, occlusions, camera motions, and UAV flying attitudes. Such a comprehensive and challenging benchmark shall be able to promote the research of UAV-based human behavior understanding, including action recognition, pose estimation, re-identification, and attribute recognition. Furthermore, we propose a fisheye-based action recognition method that mitigates the distortions in fisheye videos via learning unbounded transformations guided by flat RGB videos. Experiments show the efficacy of our method on the UAV-Human dataset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Skeleton Cloud Colorization for Unsupervised 3D Action Representation LearningSiyuan Yang, Jun Liu, Shijian Lu, Meng Hwa Er 等ICCV 2021 · 被引用 114 次
- Meta Agent Teaming Active Learning for Pose EstimationJia Gong, Zhipeng Fan, Qiuhong Ke, Hossein Rahmani 等CVPR 2022 · 被引用 53 次
- MAtch, eXpand and Improve: Unsupervised Finetuning for Zero-Shot Action Recognition with Language KnowledgeWei Lin, Leonid Karlinsky, Nina Shvetsova, Horst Possegger 等ICCV 2023 · 被引用 52 次
- Else-Net: Elastic Semantic Network for Continual Action Recognition from Skeleton DataTianjiao Li, Qiuhong Ke, Hossein Rahmani, Rui En Ho 等ICCV 2021 · 被引用 46 次
- Rotation Invariant Transformer for Recognizing Object in UAVsShuoyi Chen, Mang Ye, Bo DuACM MM 2022 · 被引用 43 次
它引用的顶会 Paper4
- Drive&Act: A Multi-Modal Dataset for Fine-Grained Driver Behavior Recognition in Autonomous VehiclesManuel Martin, Alina Roitberg, Monica Haurilet, Matthias Horne 等ICCV 2019 · 被引用 235 次
- MMAct: A Large-Scale Dataset for Cross Modal Human Action UnderstandingQuan Kong, Ziming Wu, Ziwei Deng, Martin Klinkigt 等ICCV 2019 · 被引用 108 次
- HigherHRNet: Scale-Aware Representation Learning for Bottom-Up Human Pose EstimationBowen Cheng, Bin Xiao, Jingdong Wang, Honghui Shi 等CVPR 2020
- Skeleton-Based Action Recognition With Shift Graph Convolutional NetworkKe Cheng, Yifan Zhang, Xiangyu He, Weihan Chen 等CVPR 2020
相关 Paper
- MOR-UAV: A Benchmark Dataset and Baselines for Moving Object Recognition in UAV VideosMurari Mandal, Lav Kush Kumar, Santosh Kumar VipparthiACM MM 2020 · 被引用 58 次
- Vehicle Re-Identification in Aerial Imagery: Dataset and ApproachPeng Wang, Bingliang Jiao, Lu Yang, Yifei Yang 等ICCV 2019 · 被引用 67 次
- Multi-Modal Multi-Platform Person Re-Identification: Benchmark and MethodRuiyang Ha, Songyi Jiang, Bin Li, Bikang Pan 等ICCV 2025 · 被引用 4 次
- AG-VPReID: A Challenging Large-Scale Benchmark for Aerial-Ground Video-based Person Re-IdentificationHuy Nguyen, Kien Nguyen, Akila Pemasiri, Feng Liu 等CVPR 2025
- Resource-Efficient RGBD Aerial TrackingJinyu Yang, Shang Gao, Zhe Li, Feng Zheng 等CVPR 2023
