Recurrent Homography Estimation Using Homography-Guided Image Warping and Focus Transformer
Si-Yuan Cao, Runmin Zhang, Lun Luo, Beinan Yu, Zehua Sheng, Junwei Li, Hui-Liang Shen
摘要
We propose the Recurrent homography estimation framework using Homography-guided image Warping and Focus transformer (FocusFormer), named RHWF. Both being appropriately absorbed into the recurrent framework, the homography-guided image warping progressively enhances the feature consistency and the attention-focusing mechanism in FocusFormer aggregates the intra-inter correspondence in a global→nonlocal→local manner. Thanks to the above strategies, RHWF ranks top in accuracy on a variety of datasets, including the challenging crossresolution and cross-modal ones. Meanwhile, benefiting from the recurrent framework, RHWF achieves parameter efficiency despite the transformer architecture. Compared to previous state-of-the-art approaches LocalTrans and IHN, RHWF reduces the mean average corner error (MACE) by about 70% and 38.1% on the MSCOCO dataset, while saving the parameter costs by 86.5% and 24.6%. Similar to the previous works, RHWF can also be arranged in 1-scale for efficiency and 2-scale for accuracy, with the 1scale RHWF already outperforming most of the previous methods. Source code is available at https://github. com/imdumpl78/RHWF.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Aggregating Feature Point Cloud for Depth CompletionZhu Yu, Zehua Sheng, Zili Zhou, Lun Luo 等ICCV 2023 · 被引用 42 次
- MCNet: Rethinking the Core Ingredients for Accurate and Efficient Homography EstimationHaokai Zhu, Si-Yuan Cao, Jianxin Hu, Sitong Zuo 等CVPR 2024 · 被引用 18 次
- Reconstructing the Image Stitching Pipeline: Integrating Fusion and Rectangling into a Unified Inpainting ModelZiqi Xie, Weidong Zhao, Xianhui Liu, Jian Zhao 等NeurIPS 2024 · 被引用 11 次
- Unsupervised Homography Estimation on Multimodal Image Pair via Alternating OptimizationSanghyeob Song, Jaihyun Lew, Hyemi Jang, Sungroh YoonNeurIPS 2024 · 被引用 9 次
- Semantic Ambiguity Modeling and Propagation for Fine-Grained Visual Cross View Geo-LocalizationMingtao Feng, Fenghao Tian, Jianqiao Luo, Zijie Wu 等AAAI 2025 · 被引用 4 次
它引用的顶会 Paper10
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Revisiting Stereo Depth Estimation From a Sequence-to-Sequence Perspective with TransformersZhaoshuo Li, Xingtong Liu, Nathan Drenkow, Andy S. Ding 等ICCV 2021 · 被引用 380 次
- GMFlow: Learning Optical Flow via Global MatchingHaofei Xu, Jing Zhang, Jianfei Cai, Hamid Rezatofighi 等CVPR 2022 · 被引用 353 次
- Iterative Deep Homography EstimationSi-Yuan Cao, Jianxin Hu, Ze-Hua Sheng, Hui-Liang ShenCVPR 2022 · 被引用 65 次
- Unsupervised Homography Estimation with Coplanarity-Aware GANMingbo Hong, Yuhang Lu, Nianjin Ye, Chunyu Lin 等CVPR 2022 · 被引用 62 次
相关 Paper
- Geometrized Transformer for Self-Supervised Homography EstimationJiazhen Liu, Xirong LiICCV 2023 · 被引用 27 次
- LocalTrans: A Multiscale Local Transformer Network for Cross-Resolution Homography EstimationRuizhi Shao, Gaochang Wu, Yuemei Zhou, Ying Fu 等ICCV 2021 · 被引用 57 次
- CRFT: Consistent-Recurrent Feature Flow Transformer for Cross-Modal Image RegistrationXuecong Liu, Mengzhu Ding, Zixuan Sun, Zhang Li 等CVPR 2026 · 被引用 4 次
- SSHNet: Unsupervised Cross-modal Homography Estimation via Problem Reformulation and Split OptimizationJunchen Yu, Si-Yuan Cao, Runmin Zhang, Chenghao Zhang 等CVPR 2025
- HRFormer: High-Resolution Vision Transformer for Dense PredictYuhui Yuan, Rao Fu, Lang Huang, Weihong Lin 等NeurIPS 2021 · 被引用 357 次
