Scale-aware Recognition in Satellite Images under Resource Constraints
Shreelekha Revankar, Cheng Perng Phoo, Utkarsh Mall, Bharath Hariharan, Kavita Bala
摘要
Recognition of features in satellite imagery (forests, swimming pools, etc.) depends strongly on the spatial scale of the concept and therefore the resolution of the images. This poses two challenges: Which resolution is best suited for recognizing a given concept, and where and when should the costlier higher-resolution (HR) imagery be acquired? We present a novel scheme to address these challenges by introducing three components: (1) A technique to distill knowledge from models trained on HR imagery to recognition models that operate on imagery of lower resolution (LR), (2) a sampling strategy for HR imagery based on model disagreement, and (3) an LLM-based approach for inferring concept "scale". With these components we present a system to efficiently perform scale-aware recognition in satellite imagery, improving accuracy over single-scale inference while following budget constraints. Our novel approach offers up to a 26.3% improvement over entirely HR baselines, using 76.3% fewer HR images. Resources are available on our website. Figure 1 : With these images we can see how concept scale is linked to spatial resolution. If we are seeking out a spatially large concept like forest, lower resolutions are favored (b), as higher resolutions may lack the needed context to discern between a forest (a) and a park (c). At the same time while seeking out finer concepts such as sports track, certain details can only be discerned well at higher resolutions (d) and are obscured at lower resolutions (e).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford 等ICLR 2020 · 被引用 974 次
- On the Efficacy of Knowledge DistillationJang Hyun Cho, Bharath HariharanICCV 2019 · 被引用 741 次
- SatMAE: Pre-training Transformers for Temporal and Multi-Spectral Satellite ImageryYezhen Cong, Samar Khanna, Chenlin Meng, Patrick Liu 等NeurIPS 2022 · 被引用 707 次
- A Survey on In-context LearningQingxiu Dong, Lei Li, Damai Dai, Ce Zheng 等EMNLP 2024 · 被引用 479 次
相关 Paper
- Spatial-Temporal Super-Resolution of Satellite Imagery via Conditional Pixel SynthesisYutong He, Dingjie Wang, Nicholas Lai, William Zhang 等NeurIPS 2021 · 被引用 36 次
- Efficient Poverty Mapping from High Resolution Remote Sensing ImagesKumar Ayush, Burak Uzkent, Kumar Tanmay, Marshall Burke 等AAAI 2021 · 被引用 51 次
- Look One and More: Distilling Hybrid Order Relational Knowledge for Cross-Resolution Image RecognitionShiming Ge, Kangkai Zhang, Haolin Liu, Yingying Hua 等AAAI 2020 · 被引用 30 次
- Spectrally Distilled Representations Aligned with Instruction-Augmented LLMs for Satellite ImageryMinh Kha Do, Wei Xiang, Kang Han, Di Wu 等CVPR 2026
- Foreground-Aware Relation Network for Geospatial Object Segmentation in High Spatial Resolution Remote Sensing ImageryZhuo Zheng, Yanfei Zhong, Junjue Wang, Ailong MaCVPR 2020
