Multiview Aerial Visual Recognition (MAVREC): Can Multi-View Improve Aerial Visual Perception?
Aritra Dutta, Srijan Das, Jacob Nielsen, Rajatsubhra Chakraborty, Mubarak Shah
摘要
Figure 1. Illustration of the geography-aware model using our proposed MAVREC dataset (green box) collected in the rural and urban European landscape vs. the conventional aerial object detector (blue box) pretrained only on aerial images from VisDrone [91] captured in Asia. The conventional approach fails to detect aerial objects from the MAVREC dataset precisely. In contrast, our object detector pretrained on the ground and aerial images from the MAVREC dataset contextualizes the object proposals of that specific geography and enhances the aerial visual perception, thus outperforming other object detectors pre-trained on popular ground-view dataset (MS-COCO [44]) or other aerial datasets collected from different geographies; also, see Figure 5 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- CYCLO: Cyclic Graph Transformer Approach to Multi-Object Relationship Modeling in Aerial VideosTrong-Thuan Nguyen, Pha A. Nguyen, Xin Li, Jackson David Cothren 等NeurIPS 2024 · 被引用 13 次
- Video2BEV: Transforming Drone Videos to BEVs for Video-Based Geo-LocalizationHao Ju, Shaofei Huang, Si Liu, Zhedong ZhengICCV 2025 · 被引用 5 次
- VGGT-Segmentor: Geometry-Enhanced Cross-View SegmentationYulu Gao, Bohao Zhang, Zongheng Tang, Jitong Liao 等CVPR 2026 · 被引用 3 次
- CrossVL: Complexity-Aware Feature Routing and Paired Curriculum for Cross-View Vision-Language DetectionZhipeng Liu, Chunbo LuoCVPR 2026 · 被引用 1 次
- LLAVIDAL: A Large LAnguage VIsion Model for Daily Activities of LivingDominick Reilly, Rajatsubhra Chakraborty, Arkaprava Sinha, Manish Kumar Govind 等CVPR 2025
它引用的顶会 Paper11
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- Unbiased Teacher v2: Semi-supervised Object Detection for Anchor-free and Anchor-based DetectorsYen-Cheng Liu, Chih-Yao Ma, Zsolt KiraCVPR 2022 · 被引用 124 次
- Active Teacher for Semi-Supervised Object DetectionPeng Mi, Jianghang Lin, Yiyi Zhou, Yunhang Shen 等CVPR 2022 · 被引用 83 次
- Dense Learning based Semi-Supervised Object DetectionBinghui Chen, Pengyu Li, Xiang Chen, Biao Wang 等CVPR 2022 · 被引用 80 次
- MOR-UAV: A Benchmark Dataset and Baselines for Moving Object Recognition in UAV VideosMurari Mandal, Lav Kush Kumar, Santosh Kumar VipparthiACM MM 2020 · 被引用 58 次
相关 Paper
- MODA: The First Challenging Benchmark for Multispectral Object Detection in Aerial ImagesShuaihao Han, Tingfa Xu, Peifu Liu, Jianan LiAAAI 2026 · 被引用 1 次
- UniGeoRS: A Unified Benchmark for Tri-view Geo-LocalizationXiao Liang, Huaizhi Tang, Feiyang Zhang, Shiji Yuan 等CVPR 2026
- MMGeo: Multimodal Compositional Geo-Localization for UAVsYuxiang Ji, Boyong He, Zhuoyue Tan, Liaoni WuICCV 2025 · 被引用 5 次
- GeoMIM: Towards Better 3D Knowledge Transfer via Masked Image Modeling for Multi-view 3D UnderstandingJihao Liu, Tai Wang, Boxiao Liu, Qihang Zhang 等ICCV 2023 · 被引用 22 次
- AerialVG: A Challenging Benchmark for Aerial Visual Grounding by Exploring Positional RelationsJunli Liu, Qizhi Chen, Zhigang Wang, Yiwen Tang 等ICCV 2025 · 被引用 5 次
