Gated Spatio-Temporal Attention-Guided Video Deblurring
Maitreya Suin, A. N. Rajagopalan
Abstract
Video deblurring remains a challenging task due to the complexity of spatially and temporally varying blur. Most of the existing works depend on implicit or explicit alignment for temporal information fusion, which either increases the computational cost or results in suboptimal performance due to misalignment. In this work, we investigate two key factors responsible for deblurring quality: how to fuse spatio-temporal information and from where to collect it. We propose a factorized gated spatio-temporal attention module to perform non-local operations across space and time to fully utilize the available information without depending on alignment. First, we perform spatial aggregation followed by a temporal aggregation step. Next, we adaptively distribute the global spatio-temporal information to each pixel. It shows superior performance compared to existing non-local fusion techniques while being considerably more efficient. To complement the attention module, we propose a reinforcement learning-based framework for selecting keyframes from the neighborhood with the most complementary and useful information. Moreover, our adaptive approach can increase or decrease the frame usage at inference time, depending on the user's need. Extensive experiments on multiple datasets demonstrate the superiority of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers13
- Recurrent Video Restoration Transformer with Guided Deformable AttentionJingyun Liang, Yuchen Fan, Xiaoyu Xiang, Rakesh Ranjan et al.NeurIPS 2022 · 318 citations
- Unifying Motion Deblurring and Frame Interpolation with EventsXiang Zhang, Lei YuCVPR 2022 · 85 citations
- Flow-Guided Sparse Transformer for Video DeblurringJing Lin, Yuanhao Cai, Xiaowan Hu, Haoqian Wang et al.ICML 2022 · 82 citations
- Deep Recurrent Neural Network with Multi-Scale Bi-directional Propagation for Video DeblurringChao Zhu, Hang Dong, Jinshan Pan, Boyang Liang et al.AAAI 2022 · 63 citations
- Distillation-guided Image InpaintingMaitreya Suin, Kuldeep Purohit, A. N. RajagopalanICCV 2021 · 42 citations
Builds on5
- Expectation-Maximization Attention Networks for Semantic SegmentationXia Li, Zhisheng Zhong, Jianlong Wu, Yibo Yang et al.ICCV 2019 · 639 citations
- Spatio-Temporal Filter Adaptive Network for Video DeblurringShangchen Zhou, Jiawei Zhang, Jinshan Pan, Wangmeng Zuo et al.ICCV 2019 · 225 citations
- An Efficient Framework for Dense Video CaptioningMaitreya Suin, A. N. RajagopalanAAAI 2020 · 49 citations
- Spatially-Attentive Patch-Hierarchical Network for Adaptive Motion DeblurringMaitreya Suin, Kuldeep Purohit, A. N. RajagopalanCVPR 2020
- Cascaded Deep Video Deblurring Using Temporal Sharpness PriorJinshan Pan, Haoran Bai, Jinhui TangCVPR 2020
Related papers
- Deep Discriminative Spatial and Temporal Network for Efficient Video DeblurringJinshan Pan, Boming Xu, Jiangxin Dong, Jianjun Ge et al.CVPR 2023
- ALANET: Adaptive Latent Attention Network for Joint Video Deblurring and InterpolationAkash Gupta, Abhishek Aich, Amit K. Roy-ChowdhuryACM MM 2020 · 19 citations
- Revisiting Temporal Alignment for Video RestorationKun Zhou, Wenbo Li, Liying Lu, Xiaoguang Han et al.CVPR 2022
- STRONG: Spatio-Temporal Reinforcement Learning for Cross-Modal Video Moment LocalizationDa Cao, Yawen Zeng, Meng Liu, Xiangnan He et al.ACM MM 2020 · 47 citations
- Alignment-guided Temporal Attention for Video Action RecognitionYizhou Zhao, Zhenyang Li, Xun Guo, Yan LuNeurIPS 2022 · 24 citations
