Learning to Downsample for Segmentation of Ultra-High Resolution Images
Chen Jin, Ryutaro Tanno, Thomy Mertzanidou, Eleftheria Panagiotaki, Daniel C. Alexander
Abstract
Many computer vision systems require low-cost segmentation algorithms based on deep learning, either because of the enormous size of input images or limited computational budget. Common solutions uniformly downsample the input images to meet memory constraints, assuming all pixels are equally informative. In this work, we demonstrate that this assumption can harm the segmentation performance because the segmentation difficulty varies spatially (see Figure 1 "Uniform"). We combat this problem by introducing a learnable downsampling module, which can be optimised together with the given segmentation model in an end-to-end fashion. We formulate the problem of training such downsampling module as optimisation of sampling density distributions over the input images given their low-resolution views. To defend against degenerate solutions (e.g. over-sampling trivial regions like the backgrounds), we propose a regularisation term that encourages the sampling locations to concentrate around the object boundaries. We find the downsampling module learns to sample more densely at difficult locations, thereby improving the segmentation performance (see Figure 1 "Ours"). Our experiments on benchmarks of high-resolution street view, aerial and medical images demonstrate substantial improvements in terms of efficiency-and-accuracy trade-off compared to both uniform downsampling and two recent advanced downsampling techniques. Video demos available at https://lxasqjc.github.io/learn-downsample.github.io/
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bb8ca976-833a-4854-9eb7-d3e4d75d6135Cited by top-tier papers10
- Learning to Upsample by Learning to SampleWenze Liu, Hao Lu, Hongtao Fu, Zhiguo CaoICCV 2023 · 518 citations
- Learning Strides in Convolutional Neural NetworksRachid Riad, Olivier Teboul, David Grangier, Neil ZeghidourICLR 2022 · 54 citations
- LookWhere? Efficient Visual Recognition by Learning Where to Look and What to See from Self-SupervisionAnthony Fuller, Yousef Yassin, Junfeng Wen, Tarek Ibrahim et al.NeurIPS 2025 · 7 citations
- When Visual Grounding Meets Gigapixel-Level Large-Scale Scenes: Benchmark and ApproachM. Tao, Bing Bai, Haozhe Lin, Heyuan Wang et al.CVPR 2024 · 4 citations
- Scale-Space Hypernetworks for Efficient Biomedical Image AnalysisJose Javier Gonzalez Ortiz, John V. Guttag, Adrian V. DalcaNeurIPS 2023 · 1 citation
Builds on5
- Learning to Resize Images for Computer Vision TasksHossein Talebi, Peyman MilanfarICCV 2021 · 159 citations
- Efficient Segmentation: Learning Downsampling Near Semantic BoundariesDmitrii Marin, Zijian He, Peter Vajda, Priyam Chatterjee et al.ICCV 2019 · 107 citations
- PANDA: A Gigapixel-Level Human-Centric Video DatasetXueyang Wang, Xiya Zhang, Yinheng Zhu, Yuchen Guo et al.CVPR 2020
- Progressive Semantic SegmentationChuong Huynh, Anh Tuan Tran, Khoa Luu, Minh HoaiCVPR 2021
- PointRend: Image Segmentation As RenderingAlexander Kirillov, Yuxin Wu, Kaiming He, Ross B. GirshickCVPR 2020
Related papers
- Beyond Predictive Resampling: Learning Input-Agnostic Downsampling for Efficient Aligned Vision RecognitionKai Zhao, Liting Ruan, Haoran Jiang, Xiaoqiang Zhu et al.AAAI 2026
- AutoFocusFormer: Image Segmentation off the GridZiwen Chen, Kaushik Patnaik, Shuangfei Zhai, Alvin Wan et al.CVPR 2023
- Weakly Supervised Segmentation with Point Annotations for Histopathology Images via Contrast-Based Variational ModelHongrun Zhang, Liam Burrows, Yanda Meng, Declan Sculthorpe et al.CVPR 2023
- Dense Unsupervised Learning for Video SegmentationNikita Araslanov, Simone Schaub-Meyer, Stefan RothNeurIPS 2021 · 41 citations
- Learning to Zoom and UnzoomChittesh Thavamani, Mengtian Li, Francesco Ferroni, Deva RamananCVPR 2023
