SeeClear: Semantic Distillation Enhances Pixel Condensation for Video Super-Resolution
Qi Tang, Yao Zhao, Meiqin Liu, Chao Yao
摘要
Diffusion-based Video Super-Resolution (VSR) is renowned for generating perceptually realistic videos, yet it grapples with maintaining detail consistency across frames due to stochastic fluctuations. The traditional approach of pixel-level alignment is ineffective for diffusion-processed frames because of iterative disruptions. To overcome this, we introduce SeeClear--a novel VSR framework leveraging conditional video generation, orchestrated by instance-centric and channel-wise semantic controls. This framework integrates a Semantic Distiller and a Pixel Condenser, which synergize to extract and upscale semantic details from low-resolution frames. The Instance-Centric Alignment Module (InCAM) utilizes video-clip-wise tokens to dynamically relate pixels within and across frames, enhancing coherency. Additionally, the Channel-wise Texture Aggregation Memory (CaTeGory) infuses extrinsic knowledge, capitalizing on long-standing semantic textures. Our method also innovates the blurring diffusion process with the ResShift mechanism, finely balancing between sharpness and diffusion effects. Comprehensive experiments confirm our framework's advantage over state-of-the-art diffusion-based VSR techniques. The code is available: https://github.com/Tang1705/SeeClear-NeurIPS24.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- DAM-VSR: Disentanglement of Appearance and Motion for Video Super-ResolutionZhe Kong, Le Li, Yong Zhang, Feng Gao 等SIGGRAPH 2025 · 被引用 6 次
- QD-PCQA: Quality-Aware Domain Adaptation for Point Cloud Quality AssessmentGuohua Zhang, Jian Jin, Meiqin Liu, Chao Yao 等CVPR 2026 · 被引用 3 次
- Spatial Imputation Drives Cross-Domain Alignment for EEG ClassificationHongjun Liu, Chao Yao, Yalan Zhang, Xiaokun Wang 等ACM MM 2025 · 被引用 1 次
- Proper Hölder-Kullback Dirichlet Diffusion: A Framework for High Dimensional Generative ModelingWanpeng Zhang, Yuhao Fang, Xihang Qiu, Jiarong Cheng 等NeurIPS 2025
它引用的顶会 Paper27
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Exploring CLIP for Assessing the Look and Feel of ImagesJianyi Wang, Kelvin C. K. Chan, Chen Change LoyAAAI 2023 · 被引用 1,208 次
- Maximum Likelihood Training of Score-Based Diffusion ModelsYang Song, Conor Durkan, Iain Murray, Stefano ErmonNeurIPS 2021 · 被引用 958 次
- ResShift: Efficient Diffusion Model for Image Super-resolution by Residual ShiftingZongsheng Yue, Jianyi Wang, Chen Change LoyNeurIPS 2023 · 被引用 646 次
相关 Paper
- Semantic Lens: Instance-Centric Semantic Alignment for Video Super-resolutionQi Tang, Yao Zhao, Meiqin Liu, Jian Jin 等AAAI 2024 · 被引用 10 次
- PatchVSR: Breaking Video Diffusion Resolution Limits with Patch-wise Video Super-ResolutionShian Du, Menghan Xia, Chang Liu, Xintao Wang 等CVPR 2025
- FlashVSR: Towards Real-time Diffusion-Based Streaming Video Super ResolutionJunhao Zhuang, Shi Guo, Xin Cai, Xiaohui Li 等CVPR 2026 · 被引用 42 次
- CoSeR: Bridging Image and Language for Cognitive Super-ResolutionHaoze Sun, Wenbo Li, Jianzhuang Liu, Haoyu Chen 等CVPR 2024 · 被引用 43 次
- Rethinking Diffusion Model-Based Video Super-Resolution: Leveraging Dense Guidance from Aligned FeaturesJingyi Xu, Meisong Zheng, Ying Chen, Minglang Qiao 等CVPR 2026 · 被引用 1 次
