High-Resolution Optical Flow from 1D Attention and Correlation
Haofei Xu, Jiaolong Yang, Jianfei Cai, Juyong Zhang, Xin Tong
摘要
Optical flow is inherently a 2D search problem, and thus the computational complexity grows quadratically with respect to the search window, making large displacements matching infeasible for high-resolution images. In this paper, we take inspiration from Transformers and propose a new method for high-resolution optical flow estimation with significantly less computation. Specifically, a 1D attention operation is first applied in the vertical direction of the target image, and then a simple 1D correlation in the horizontal direction of the attended image is able to achieve 2D correspondence modeling effect. The directions of attention and correlation can also be exchanged, resulting in two 3D cost volumes that are concatenated for optical flow estimation. The novel 1D formulation empowers our method to scale to very high-resolution input images while maintaining competitive performance. Extensive experiments on Sintel, KITTI and real-world 4K (2160 × 3840) resolution images demonstrated the effectiveness and superiority of our proposed method. Code and models are available at https://github.com/haofeixu/flow1d .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper35
- GMFlow: Learning Optical Flow via Global MatchingHaofei Xu, Jing Zhang, Jianfei Cai, Hamid Rezatofighi 等CVPR 2022 · 被引用 353 次
- VideoFlow: Exploiting Temporal Cues for Multi-frame Optical Flow EstimationXiaoyu Shi, Zhaoyang Huang, Weikang Bian, Dasong Li 等ICCV 2023 · 被引用 112 次
- SKFlow: Learning Optical Flow with Super KernelsShangkun Sun, Yuanqi Chen, Yu Zhu, Guodong Guo 等NeurIPS 2022 · 被引用 97 次
- Learning Optical Flow with Kernel Patch AttentionAo Luo, Fan Yang, Xin Li, Shuaicheng LiuCVPR 2022 · 被引用 63 次
- Implicit Warping for Animation with Image SetsArun Mallya, Ting-Chun Wang, Ming-Yu LiuNeurIPS 2022 · 被引用 62 次
它引用的顶会 Paper3
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang 等ICCV 2019 · 被引用 2,972 次
- MaskFlownet: Asymmetric Feature Matching With Learnable Occlusion MaskShengyu Zhao, Yilun Sheng, Yue Dong, Eric I-Chao Chang 等CVPR 2020
- AANet: Adaptive Aggregation Network for Efficient Stereo MatchingHaofei Xu, Juyong ZhangCVPR 2020
相关 Paper
- TransFlow: Transformer as Flow LearnerYawen Lu, Qifan Wang, Siqi Ma, Tong Geng 等CVPR 2023
- CRAFT: Cross-Attentional Flow Transformer for Robust Optical FlowXiuchao Sui, Shaohua Li, Xue Geng, Yan Wu 等CVPR 2022 · 被引用 114 次
- DIP: Deep Inverse Patchmatch for High-Resolution Optical FlowZihua Zheng, Ni Nie, Zhi Ling, Pengfei Xiong 等CVPR 2022 · 被引用 48 次
- CoWTracker: Tracking by Warping instead of CorrelationZihang Lai, Eldar Insafutdinov, Edgar Sucar, Andrea VedaldiCVPR 2026 · 被引用 12 次
- Video Frame Interpolation with Flow TransformerPan Gao, Haoyue Tian, Jie QinACM MM 2023 · 被引用 4 次
