ALANET: Adaptive Latent Attention Network for Joint Video Deblurring and Interpolation
Akash Gupta, Abhishek Aich, Amit K. Roy-Chowdhury
摘要
Existing works address the problem of generating high frame-rate sharp videos by separately learning the frame deblurring and frame interpolation modules. Most of these approaches have a strong prior assumption that all the input frames are blurry whereas in a real-world setting, the quality of frames varies. Moreover, such approaches are trained to perform either of the two tasks - deblurring or interpolation - in isolation, while many practical situations call for both. Different from these works, we address a more realistic problem of high frame-rate sharp video synthesis with no prior assumption that input is always blurry. We introduce a novel architecture, Adaptive Latent Attention Network (ALANET), which synthesizes sharp high frame-rate videos with no prior knowledge of input frames being blurry or not, thereby performing the task of both deblurring and interpolation. We hypothesize that information from the latent representation of the consecutive frames can be utilized to generate optimized representations for both frame deblurring and frame interpolation. Specifically, we employ combination of self-attention and cross-attention module between consecutive frames in the latent space to generate optimized representation for each frame. The optimized representation learnt using these attention modules help the model to generate and interpolate sharp frames. Extensive experiments on standard datasets demonstrate that our method performs favorably against various state-of-the-art approaches, even though we tackle a much more difficult problem. The project page is available at https://agupt013.github.io/ALANET.html.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Spatio-Temporal Representation Factorization for Video-based Person Re-IdentificationAbhishek Aich, Meng Zheng, Srikrishna Karanam, Terrence Chen 等ICCV 2021 · 被引用 86 次
- Adversarial Attacks on Black Box Video Classifiers: Leveraging the Power of Geometric TransformationsShasha Li, Abhishek Aich, Shitong Zhu, M. Salman Asif 等NeurIPS 2021 · 被引用 50 次
- GAMA: Generative Adversarial Multi-Object Scene AttacksAbhishek Aich, Calvin-Khang Ta, Akash Gupta, Chengyu Song 等NeurIPS 2022 · 被引用 26 次
- FMA-Net: Flow-Guided Dynamic Filtering and Iterative Feature Refinement with Multi-Attention for Joint Video Super-Resolution and DeblurringGeunhyuk Youk, Jihyong Oh, Munchurl KimCVPR 2024 · 被引用 16 次
- Ada-VSR: Adaptive Video Super-Resolution with Meta-LearningAkash Gupta, Padmaja Jonnalagedda, Bir Bhanu, Amit K. Roy-ChowdhuryACM MM 2021 · 被引用 9 次
它引用的顶会 Paper4
- Channel Attention Is All You Need for Video Frame InterpolationMyungsub Choi, Heewon Kim, Bohyung Han, Ning Xu 等AAAI 2020 · 被引用 362 次
- Spatio-Temporal Filter Adaptive Network for Video DeblurringShangchen Zhou, Jiawei Zhang, Jinshan Pan, Wangmeng Zuo 等ICCV 2019 · 被引用 225 次
- Non-Adversarial Video Synthesis with Learned PriorsAbhishek Aich, Akash Gupta, Rameswar Panda, Rakib Hyder 等CVPR 2020
- Blurry Video Frame InterpolationWang Shen, Wenbo Bao, Guangtao Zhai, Li Chen 等CVPR 2020
相关 Paper
- Event-Based Frame Interpolation with Ad-hoc DeblurringLei Sun, Christos Sakaridis, Jingyun Liang, Peng Sun 等CVPR 2023
- Space-Time-Aware Multi-Resolution Video EnhancementMuhammad Haris, Greg Shakhnarovich, Norimichi UkitaCVPR 2020
- Unifying Motion Deblurring and Frame Interpolation with EventsXiang Zhang, Lei YuCVPR 2022 · 被引用 85 次
- Time-Specialized Event-Image Alignment for Blur-to-Video DecompositionZhijing Sun, Senyan Xu, Ruixuan Jiang, Kean Liu 等CVPR 2026
- Joint Video Multi-Frame Interpolation and Deblurring under Unknown Exposure TimeWei Shang, Dongwei Ren, Yi Yang, Hongzhi Zhang 等CVPR 2023
