Learning When and Where to Zoom With Deep Reinforcement Learning
Burak Uzkent, Stefano Ermon
Abstract
While high resolution images contain semantically more useful information than their lower resolution counterparts, processing them is computationally more expensive, and in some applications, e.g. remote sensing, they can be much more expensive to acquire. For these reasons, it is desirable to develop an automatic method to selectively use high resolution data when necessary while maintaining accuracy and reducing acquisition/run-time cost. In this direction, we propose PatchDrop a reinforcement learning approach to dynamically identify when and where to use/acquire high resolution data conditioned on the paired, cheap, low resolution images. We conduct experiments on CIFAR10, CI-FAR100, ImageNet and fMoW datasets where we use significantly less high resolution data while maintaining similar accuracy to models which use full high resolution images.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3e97b32e-43e1-45b6-b534-30f56bd33503Cited by top-tier papers22
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- Dynamic Resolution NetworkMingjian Zhu, Kai Han, Enhua Wu, Qiulin Zhang et al.NeurIPS 2021 · 71 citations
- Efficient Poverty Mapping from High Resolution Remote Sensing ImagesKumar Ayush, Burak Uzkent, Kumar Tanmay, Marshall Burke et al.AAAI 2021 · 51 citations
- FOVEA: Foveated Image Magnification for Autonomous NavigationChittesh Thavamani, Mengtian Li, Nicolas Cebron, Deva RamananICCV 2021 · 45 citations
- Hard-Attention for Scalable Image ClassificationAthanasios Papadopoulos, Pawel Korus, Nasir D. MemonNeurIPS 2021 · 38 citations
Related papers
- Glance and Focus: a Dynamic Approach to Reducing Spatial Redundancy in Image ClassificationYulin Wang, Kangchen Lv, Rui Huang, Shiji Song et al.NeurIPS 2020 · 179 citations
- Spatial-Temporal Super-Resolution of Satellite Imagery via Conditional Pixel SynthesisYutong He, Dingjie Wang, Nicholas Lai, William Zhang et al.NeurIPS 2021 · 36 citations
- Differentiable Patch Selection for Image RecognitionJean-Baptiste Cordonnier, Aravindh Mahendran, Alexey Dosovitskiy, Dirk Weissenborn et al.CVPR 2021
- SelectAugment: Hierarchical Deterministic Sample Selection for Data AugmentationShiqi Lin, Zhizheng Zhang, Xin Li, Zhibo ChenAAAI 2023 · 13 citations
- Self-Play Reinforcement Learning for Fast Image RetargetingNobukatsu Kajiura, Satoshi Kosugi, Xueting Wang, Toshihiko YamasakiACM MM 2020 · 21 citations
