Gated Spatio-Temporal Attention-Guided Video Deblurring
Maitreya Suin, A. N. Rajagopalan
摘要
Video deblurring remains a challenging task due to the complexity of spatially and temporally varying blur. Most of the existing works depend on implicit or explicit alignment for temporal information fusion, which either increases the computational cost or results in suboptimal performance due to misalignment. In this work, we investigate two key factors responsible for deblurring quality: how to fuse spatio-temporal information and from where to collect it. We propose a factorized gated spatio-temporal attention module to perform non-local operations across space and time to fully utilize the available information without depending on alignment. First, we perform spatial aggregation followed by a temporal aggregation step. Next, we adaptively distribute the global spatio-temporal information to each pixel. It shows superior performance compared to existing non-local fusion techniques while being considerably more efficient. To complement the attention module, we propose a reinforcement learning-based framework for selecting keyframes from the neighborhood with the most complementary and useful information. Moreover, our adaptive approach can increase or decrease the frame usage at inference time, depending on the user's need. Extensive experiments on multiple datasets demonstrate the superiority of our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Recurrent Video Restoration Transformer with Guided Deformable AttentionJingyun Liang, Yuchen Fan, Xiaoyu Xiang, Rakesh Ranjan 等NeurIPS 2022 · 被引用 318 次
- Unifying Motion Deblurring and Frame Interpolation with EventsXiang Zhang, Lei YuCVPR 2022 · 被引用 85 次
- Flow-Guided Sparse Transformer for Video DeblurringJing Lin, Yuanhao Cai, Xiaowan Hu, Haoqian Wang 等ICML 2022 · 被引用 82 次
- Deep Recurrent Neural Network with Multi-Scale Bi-directional Propagation for Video DeblurringChao Zhu, Hang Dong, Jinshan Pan, Boyang Liang 等AAAI 2022 · 被引用 63 次
- Distillation-guided Image InpaintingMaitreya Suin, Kuldeep Purohit, A. N. RajagopalanICCV 2021 · 被引用 42 次
它引用的顶会 Paper5
- Expectation-Maximization Attention Networks for Semantic SegmentationXia Li, Zhisheng Zhong, Jianlong Wu, Yibo Yang 等ICCV 2019 · 被引用 639 次
- Spatio-Temporal Filter Adaptive Network for Video DeblurringShangchen Zhou, Jiawei Zhang, Jinshan Pan, Wangmeng Zuo 等ICCV 2019 · 被引用 225 次
- An Efficient Framework for Dense Video CaptioningMaitreya Suin, A. N. RajagopalanAAAI 2020 · 被引用 49 次
- Spatially-Attentive Patch-Hierarchical Network for Adaptive Motion DeblurringMaitreya Suin, Kuldeep Purohit, A. N. RajagopalanCVPR 2020
- Cascaded Deep Video Deblurring Using Temporal Sharpness PriorJinshan Pan, Haoran Bai, Jinhui TangCVPR 2020
相关 Paper
- Deep Discriminative Spatial and Temporal Network for Efficient Video DeblurringJinshan Pan, Boming Xu, Jiangxin Dong, Jianjun Ge 等CVPR 2023
- ALANET: Adaptive Latent Attention Network for Joint Video Deblurring and InterpolationAkash Gupta, Abhishek Aich, Amit K. Roy-ChowdhuryACM MM 2020 · 被引用 19 次
- Revisiting Temporal Alignment for Video RestorationKun Zhou, Wenbo Li, Liying Lu, Xiaoguang Han 等CVPR 2022
- STRONG: Spatio-Temporal Reinforcement Learning for Cross-Modal Video Moment LocalizationDa Cao, Yawen Zeng, Meng Liu, Xiangnan He 等ACM MM 2020 · 被引用 47 次
- Alignment-guided Temporal Attention for Video Action RecognitionYizhou Zhao, Zhenyang Li, Xun Guo, Yan LuNeurIPS 2022 · 被引用 24 次
