VideoGigaGAN: Towards Detail-rich Video Super-Resolution
Yiran Xu, Taesung Park, Richard Zhang, Yang Zhou, Eli Shechtman, Feng Liu, Jia-Bin Huang, Difan Liu
摘要
Video super-resolution (VSR) models achieve temporal consistency but often produce blurrier results than their image-based counterparts due to limited generative capacity. This prompts the question: can we adapt a generative image upsampler for VSR while preserving temporal consistency? We introduce VideoGigaGAN, a new generative VSR model that combines high-frequency detail with temporal stability, building on the large-scale GigaGAN image upsampler. Simple adaptations of GigaGAN for VSR led to flickering issues, so we propose techniques to enhance temporal consistency. We validate the effectiveness of VideoGigaGAN by comparing it with state-of-the-art VSR models on public datasets and showcasing video results with 8× upsampling.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- DAM-VSR: Disentanglement of Appearance and Motion for Video Super-ResolutionZhe Kong, Le Li, Yong Zhang, Feng Gao 等SIGGRAPH 2025 · 被引用 6 次
- Improved Adversarial Diffusion Compression for Real-World Video Super-ResolutionBin Chen, Weiqi Li, Shijie Zhao, Xuanyu Zhang 等ICLR 2026 · 被引用 5 次
- DUO-VSR: Dual-Stream Distillation for One-Step Video Super-ResolutionZhengyao Lv, Menghan Xia, Xintao Wang, Kwan-Yee K. WongCVPR 2026 · 被引用 4 次
- SR3R: Rethinking Super-Resolution 3D Reconstruction With Feed-Forward Gaussian SplattingXiang Feng, Xiangbo Wang, Tieshi Zhong, Chengkai Wang 等CVPR 2026 · 被引用 3 次
- DC-VSR: Spatially and Temporally Consistent Video Super-Resolution with Video Diffusion PriorJanghyeok Han, Gyujin Sim, Geonung Kim, Hyun-Seung Lee 等SIGGRAPH 2025 · 被引用 2 次
它引用的顶会 Paper31
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan 等NeurIPS 2022 · 被引用 2,948 次
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen 等NeurIPS 2021 · 被引用 2,126 次
- Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video GenerationJay Zhangjie Wu, Yixiao Ge, Xintao Wang, Stan Weixian Lei 等ICCV 2023 · 被引用 1,113 次
相关 Paper
- Star: Spatial-Temporal Augmentation with Text-to-Video Models for Real-World Video Super-ResolutionRui Xie, Yinhong Liu, Penghao Zhou, Chen Zhao 等ICCV 2025 · 被引用 11 次
- Temporal Inconsistency Guidance for Super-resolution Video Quality AssessmentYixiao Li, Xiaoyuan Yang, Weide Liu, Xin Jin 等AAAI 2026 · 被引用 3 次
- Scaling up GANs for Text-to-Image SynthesisMinguk Kang, Jun-Yan Zhu, Richard Zhang, Jaesik Park 等CVPR 2023
- VideoVAE+: Large Motion Video Autoencoding with Cross-Modal Video VAEYazhou Xing, Yang Fei, Yingqing He, Jingye Chen 等ICCV 2025 · 被引用 2 次
- PatchVSR: Breaking Video Diffusion Resolution Limits with Patch-wise Video Super-ResolutionShian Du, Menghan Xia, Chang Liu, Xintao Wang 等CVPR 2025
