Learning Optical Flow with Kernel Patch Attention
Ao Luo, Fan Yang, Xin Li, Shuaicheng Liu
摘要
Optical flow is a fundamental method used for quantitative motion estimation on the image plane. In the deep learning era, most works treat it as a task of 'matching of features', learning to pull matched pixels as close as possible in feature space and vice versa. However, spatial affinity (smoothness constraint), another important component for motion understanding, has been largely overlooked. In this paper, we introduce a novel approach, called kernel patch attention (KPA), to better resolve the ambiguity in dense matching by explicitly taking the local context relations into consideration. Our KPA operates on each local patch, and learns to mine the context affinities for better inferring the flow fields. It can be plugged into contemporary optical flow architecture and empower the model to conduct comprehensive motion analysis with both feature similarities and spatial relations. On Sintel dataset, the proposed KPA-Flow achieves the best performance with EPE of 1.35 on clean pass and 2.36 on final pass, and it sets a new record of 4.60% in F1-all on KITTI-15 benchmark. Code is publicly available at https://github.com/megvii- research/KPAFlow.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper32
- WAFT: Warping-Alone Field Transforms for Optical FlowYihan Wang, Jia DengICLR 2026 · 被引用 36 次
- GAFlow: Incorporating Gaussian Attention into Optical FlowAo Luo, Fan Yang, Xin Li, Lang Nie 等ICCV 2023 · 被引用 35 次
- AccFlow: Backward Accumulation for Long-Range Optical FlowGuangyang Wu, Xiaohong Liu, Kunming Luo, Xi Liu 等ICCV 2023 · 被引用 33 次
- Recurrent Partial Kernel Network for Efficient Optical Flow EstimationHenrique Morimitsu, Xiaobin Zhu, Xiangyang Ji, Xu-Cheng YinAAAI 2024 · 被引用 29 次
- Learning Optical Flow from Event Camera with Rendered DatasetXinglong Luo, Kunming Luo, Ao Luo, Zhengning Wang 等ICCV 2023 · 被引用 28 次
它引用的顶会 Paper16
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui 等ICCV 2019 · 被引用 3,193 次
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang 等ICCV 2019 · 被引用 2,972 次
- Asymmetric Non-Local Neural Networks for Semantic SegmentationZhen Zhu, Mengdu Xu, Song Bai, Tengteng Huang 等ICCV 2019 · 被引用 694 次
- Learning to Estimate Hidden Motions with Global Motion AggregationShihao Jiang, Dylan Campbell, Yao Lu, Hongdong Li 等ICCV 2021 · 被引用 402 次
相关 Paper
- TransFlow: Transformer as Flow LearnerYawen Lu, Qifan Wang, Siqi Ma, Tong Geng 等CVPR 2023
- Learning Optical Flow with Adaptive Graph ReasoningAo Luo, Fan Yang, Kunming Luo, Xin Li 等AAAI 2022 · 被引用 73 次
- DIP: Deep Inverse Patchmatch for High-Resolution Optical FlowZihua Zheng, Ni Nie, Zhi Ling, Pengfei Xiong 等CVPR 2022 · 被引用 48 次
- Rethinking Optical Flow from Geometric Matching Consistent PerspectiveQiaole Dong, Chenjie Cao, Yanwei FuCVPR 2023
- CRAFT: Cross-Attentional Flow Transformer for Robust Optical FlowXiuchao Sui, Shaohua Li, Xue Geng, Yan Wu 等CVPR 2022 · 被引用 114 次
