Self-Conditioned Probabilistic Learning of Video Rescaling
Yuan Tian, Guo Lu, Xiongkuo Min, Zhaohui Che, Guangtao Zhai, Guodong Guo, Zhiyong Gao
摘要
Bicubic downscaling is a prevalent technique used to reduce the video storage burden or to accelerate the downstream processing speed. However, the inverse upscaling step is non-trivial, and the downscaled video may also deteriorate the performance of downstream tasks. In this paper, we propose a self-conditioned probabilistic framework for video rescaling to learn the paired downscaling and upscaling procedures simultaneously. During the training, we decrease the entropy of the information lost in the downscaling by maximizing its probability conditioned on the strong spatial-temporal prior information within the downscaled video. After optimization, the downscaled video by our framework preserves more meaningful information, which is beneficial for both the upscaling step and the downstream tasks, e.g., video action recognition task. We further extend the framework to a lossy video compression system, in which a gradient estimator for non-differential industrial lossy codecs is proposed for the end-to-end training of the whole system. Extensive experimental results demonstrate the superiority of our approach on video rescaling, video compression, and efficient action recognition tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Non-Semantics Suppressed Mask Learning for Unsupervised Video Semantic CompressionYuan Tian, Guo Lu, Guangtao Zhai, Zhiyong GaoICCV 2023 · 被引用 29 次
- Medical Manifestation-Aware De-IdentificationYuan Tian, Shuo Wang, Guangtao ZhaiAAAI 2025 · 被引用 7 次
- Semantics Versus Identity: A Divide-and-Conquer Approach Towards Adjustable Medical Image De-IdentificationYuan Tian, Shuo Wang, Rongzhao Zhang, Zijian Chen 等ICCV 2025 · 被引用 3 次
- Continuous Space-Time Video Resampling with Invertible Motion SteganographyYuantong Zhang, Zhenzhong ChenCVPR 2025
- Task-Aware Encoder Control for Deep Video CompressionXingtong Ge, Jixiang Luo, Xinjie Zhang, Tongda Xu 等CVPR 2024
它引用的顶会 Paper10
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 被引用 4,104 次
- Progressive Fusion Video Super-Resolution Network via Exploiting Non-Local Spatio-Temporal CorrelationsPeng Yi, Zhongyuan Wang, Kui Jiang, Junjun Jiang 等ICCV 2019 · 被引用 309 次
- Video Compression With Rate-Distortion AutoencodersAmirHossein Habibian, Ties van Rozendaal, Jakub M. Tomczak, Taco CohenICCV 2019 · 被引用 233 次
- Neural Inter-Frame Compression for Video CodingAbdelaziz Djelouah, Joaquim Campos, Simone Schaub-Meyer, Christopher SchroersICCV 2019 · 被引用 207 次
- DeHiB: Deep Hidden Backdoor Attack on Semi-supervised Learning via Adversarial PerturbationZhicong Yan, Gaolei Li, Yuan Tian, Jun Wu 等AAAI 2021 · 被引用 43 次
相关 Paper
- Self-Asymmetric Invertible Network for Compression-Aware Image RescalingJinhai Yang, Mengxi Guo, Shijie Zhao, Junlin Li 等AAAI 2023 · 被引用 11 次
- Deep Hierarchical Video CompressionMing Lu, Zhihao Duan, Fengqing Zhu, Zhan MaAAAI 2024 · 被引用 19 次
- Plug-and-Play Versatile Compressed Video EnhancementHuimin Zeng, Jiacheng Li, Zhiwei XiongCVPR 2025
- Video Rescaling Networks With Joint Optimization Strategies for Downscaling and UpscalingYan-Cheng Huang, Yi-Hsin Chen, Cheng-You Lu, Hui-Po Wang 等CVPR 2021
- Timestep-Aware Diffusion Model for Extreme Image RescalingCe Wang, Zhenyu Hu, Wanjie Sun, Zhenzhong ChenICCV 2025 · 被引用 4 次
