Linearized Multi-Sampling for Differentiable Image Transformation
Wei Jiang, Weiwei Sun, Andrea Tagliasacchi, Eduard Trulls, Kwang Moo Yi
摘要
We propose a novel image sampling method for differentiable image transformation in deep neural networks. The sampling schemes currently used in deep learning, such as Spatial Transformer Networks, rely on bilinear interpolation, which performs poorly under severe scale changes, and more importantly, results in poor gradient propagation. This is due to their strict reliance on direct neighbors. Instead, we propose to generate random auxiliary samples in the vicinity of each pixel in the sampled image, and create a linear approximation with their intensity values. We then use this approximation as a differentiable formula for the transformed image. We demonstrate that our approach produces more representative gradients with a wider basin of convergence for image alignment, which leads to considerable performance improvements when training networks for registration and classification tasks. This is not only true under large downsampling, but also when there are no scale changes. We compare our approach with multi-scale sampling and show that we outperform it. We then demonstrate that our improvements to the sampler are compatible with other tangential improvements to Spatial Transformer Networks and that it further improves their performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- COTR: Correspondence Transformer for Matching Across ImagesWei Jiang, Eduard Trulls, Jan Hosang, Andrea Tagliasacchi 等ICCV 2021 · 被引用 318 次
- Context-aware Attentional Pooling (CAP) for Fine-grained Visual ClassificationArdhendu Behera, Zachary Wharton, Pradeep R. P. G. Hewage, Asish BeraAAAI 2021 · 被引用 142 次
- PointMBF: A Multi-scale Bidirectional Fusion Network for Unsupervised RGB-D Point Cloud RegistrationMingzhi Yuan, Kexue Fu, Zhihao Li, Yucong Meng 等ICCV 2023 · 被引用 29 次
- Improved Monocular Depth Prediction Using Distance Transform Over Pre-semantic Contours with Self-supervised Neural NetworksMarwane Hariat, Antoine Manzanera, David FilliatCVPR 2025
- Deep Image Spatial Transformation for Person Image GenerationYurui Ren, Xiaoming Yu, Junming Chen, Thomas H. Li 等CVPR 2020
它引用的顶会 Paper1
相关 Paper
- ICON: Learning Regular Maps Through Inverse ConsistencyThomas Hastings Greer, Roland Kwitt, François-Xavier Vialard, Marc NiethammerICCV 2021 · 被引用 38 次
- Learning Steerable Function for Efficient Image ResamplingJiacheng Li, Chang Chen, Wei Huang, Zhiqiang Lang 等CVPR 2023
- Correction Filter for Single Image Super-Resolution: Robustifying Off-the-Shelf Deep Super-ResolversShady Abu Hussein, Tom Tirer, Raja GiryesCVPR 2020
- NODEO: A Neural Ordinary Differential Equation Based Optimization Framework for Deformable Image RegistrationYifan Wu, Tom Z. Jiahao, Jiancong Wang, Paul A. Yushkevich 等CVPR 2022 · 被引用 32 次
- Generalized Differentiable RANSACTong Wei, Yash Patel, Alexander Shekhovtsov, Jirí Matas 等ICCV 2023 · 被引用 43 次
