Efficient Classification of Very Large Images with Tiny Objects
Fanjie Kong, Ricardo Henao
摘要
An increasing number of applications in computer vision, specially, in medical imaging and remote sensing, become challenging when the goal is to classify very large images with tiny informative objects. Specifically, these classification tasks face two key challenges: i) the size of the input image is usually in the order of mega- or giga-pixels, however, existing deep architectures do not easily operate on such big images due to memory constraints, consequently, we seek a memory-efficient method to process these images; and ii) only a very small fraction of the input images are informative of the label of interest, resulting in low region of interest (ROI) to image ratio. However, most of the current convolutional neural networks (CNNs) are designed for image classification datasets that have relatively large ROIs and small image sizes (sub-megapixel). Existing approaches have addressed these two challenges in isolation. We present an end-to-end CNN model termed Zoom-In network that leverages hierarchical attention sampling for classification of large images with tiny objects using a single GPU. We evaluate our method on four large-image histopathology, road-scene and satellite imaging datasets, and one gigapixel pathology dataset. Experimental results show that our model achieves higher accuracy than existing methods while requiring less memory resources.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Few-Shot Class-Incremental Learning for Named Entity RecognitionRui Wang, Tong Yu, Handong Zhao, Sungchul Kim 等ACL 2022 · 被引用 26 次
- Visual Language Pretrained Multiple Instance Zero-Shot Transfer for Histopathology ImagesMing Y. Lu, Bowen Chen, Andrew Zhang, Drew F. K. Williamson 等CVPR 2023
- No Pains, More Gains: Recycling Sub-Salient Patches for Efficient High-Resolution Image RecognitionRong Qin, Xin Liu, Xingyu Liu, Jiaxuan Liu 等CVPR 2025
它引用的顶会 Paper4
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna 等NeurIPS 2020 · 被引用 7,049 次
- Hard-Attention for Scalable Image ClassificationAthanasios Papadopoulos, Pawel Korus, Nasir D. MemonNeurIPS 2021 · 被引用 38 次
- Learning When and Where to Zoom With Deep Reinforcement LearningBurak Uzkent, Stefano ErmonCVPR 2020
- Differentiable Patch Selection for Image RecognitionJean-Baptiste Cordonnier, Aravindh Mahendran, Alexey Dosovitskiy, Dirk Weissenborn 等CVPR 2021
相关 Paper
- Sequential Attention-based Sampling for Histopathological AnalysisTarun Gogisetty, Naman Malpani, Gugan Thoppe, Sridharan DevarajanNeurIPS 2025 · 被引用 2 次
- Multi-Stage Pathological Image Classification Using Semantic SegmentationShusuke Takahama, Yusuke Kurose, Yusuke Mukuta, Hiroyuki Abe 等ICCV 2019 · 被引用 53 次
- ThumbNet: One Thumbnail Image Contains All You Need for RecognitionChen Zhao, Bernard GhanemACM MM 2020 · 被引用 14 次
- Recurrent Networks for Guided Multi-Attention ClassificationXin Dai, Xiangnan Kong, Tian Guo, John Boaz Lee 等KDD 2020 · 被引用 5 次
- Bridging Local Inductive Bias and Long-Range Dependencies With Pixel-Mamba for End-To-End Whole Slide Image AnalysisZhongwei Qiu, Hanqing Chao, Tiancheng Lin, Wanxing Chang 等ICCV 2025 · 被引用 1 次
