Learning to Upsample by Learning to Sample
Wenze Liu, Hao Lu, Hongtao Fu, Zhiguo Cao
摘要
We present DySample, an ultra-lightweight and effective dynamic upsampler. While impressive performance gains have been witnessed from recent kernel-based dynamic upsamplers such as CARAFE, FADE, and SAPA, they introduce much workload, mostly due to the time-consuming dynamic convolution and the additional sub-network used to generate dynamic kernels. Further, the need for high-res feature guidance of FADE and SAPA somehow limits their application scenarios. To address these concerns, we bypass dynamic convolution and formulate upsampling from the perspective of point sampling, which is more resource-efficient and can be easily implemented with the standard built-in function in PyTorch. We first showcase a naive design, and then demonstrate how to strengthen its upsampling behavior step by step towards our new upsampler, DySample. Compared with former kernel-based dynamic upsamplers, DySample requires no customized CUDA package and has much fewer parameters, FLOPs, GPU memory, and latency. Besides the light-weight characteristics, DySample outperforms other upsamplers across five dense prediction tasks, including semantic segmentation, object detection, instance segmentation, panoptic segmentation, and monocular depth estimation. Code is available at https://github.com/tiny-smart/dysample.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- FeatUp: A Model-Agnostic Framework for Features at Any ResolutionStephanie Fu, Mark Hamilton, Laura E. Brandt, Axel Feldmann 等ICLR 2024 · 被引用 117 次
- AnyUp: Universal Feature UpsamplingThomas Wimmer, Prune Truong, Marie-Julie Rakotosaona, Michael Oechsle 等ICLR 2026 · 被引用 29 次
- JAFAR: Jack up Any Feature at Any ResolutionPaul Couairon, Loïck Chambon, Louis Serrano, Jean-Emmanuel Haugeard 等NeurIPS 2025 · 被引用 26 次
- NAF: Zero-Shot Feature Upsampling via Neighborhood Attention FilteringLoïck Chambon, Paul Couairon, Éloi Zablocki, Alexandre Boulch 等CVPR 2026 · 被引用 6 次
- DenseSR: Image Shadow Removal as Dense PredictionYu-Fan Lin, Chia-Ming Lee, Chih-Chung HsuACM MM 2025 · 被引用 4 次
它引用的顶会 Paper15
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 被引用 6,042 次
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
相关 Paper
- CARAFE: Content-Aware ReAssembly of FEaturesJiaqi Wang, Kai Chen, Rui Xu, Ziwei Liu 等ICCV 2019 · 被引用 842 次
- SAPA: Similarity-Aware Point Affiliation for Feature UpsamplingHao Lu, Wenze Liu, Zixuan Ye, Hongtao Fu 等NeurIPS 2022 · 被引用 96 次
- LDA-AQU: Adaptive Query-guided Upsampling via Local Deformable AttentionZewen Du, Zhenjiang Hu, Guiyu Zhao, Ying Jin 等ACM MM 2024 · 被引用 4 次
- Dynamic Sampling Network for Semantic SegmentationBin Fu, Junjun He, Zhengfu Zhang, Yu QiaoAAAI 2020 · 被引用 6 次
- You Only Segment Once: Towards Real-Time Panoptic SegmentationJie Hu, Linyan Huang, Tianhe Ren, Shengchuan Zhang 等CVPR 2023
