Overfitting the Data: Compact Neural Video Delivery via Content-aware Feature Modulation
Jiaming Liu, Ming Lu, Kaixin Chen, Xiaoqi Li, Shizun Wang, Zhaoqing Wang, Enhua Wu, Yurong Chen, Chuang Zhang, Ming Wu
Abstract
Internet video delivery has undergone a tremendous explosion of growth over the past few years. However, the quality of video delivery system greatly depends on the Internet bandwidth. Deep Neural Networks (DNNs) are utilized to improve the quality of video delivery recently. These methods divide a video into chunks, and stream LR video chunks and corresponding content-aware models to the client. The client runs the inference of models to super-resolve the LR chunks. Consequently, a large number of models are streamed in order to deliver a video. In this paper, we first carefully study the relation between models of different chunks, then we tactfully design a joint training framework along with the Content-aware Feature Modulation (CaFM) layer to compress these models for neural video delivery. With our method, each video chunk only requires less than 1% of original parameters to be streamed, achieving even better SR performance. We conduct extensive experiments across various SR backbones, video time length, and scaling factors to demonstrate the advantages of our method. Besides, our method can be also viewed as a new approach of video coding. Our primary experiments achieve better video quality compared with the commercial H.264 and H.265 standard under the same storage cost, showing the great potential of the proposed method. Code is available at: https: //github.com/Neural-video-delivery/ CaFM-Pytorch-ICCV2021 * Equal Contribution. † This work was done when Jiaming Liu was an intern at Intel Labs China supervised by Ming Lu ‡ Chuang Zhang is responsible for correspondence.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Efficient Deweahter Mixture-of-Experts with Uncertainty-Aware Feature-Wise Linear ModulationRongyu Zhang, Yulin Luo, Jiaming Liu, Huanrui Yang et al.AAAI 2024 · 30 citations
- Dynamic Low-Rank Instance Adaptation for Universal Neural Image CompressionYue Lv, Jinxi Xiang, Jun Zhang, Wenming Yang et al.ACM MM 2023 · 22 citations
- On the Robustness of Neural-Enhanced Video Streaming against Adversarial AttacksQihua Zhou, Jingcai Guo, Song Guo, Ruibin Li et al.AAAI 2024 · 7 citations
- DeNC: Unleash Neural Codecs in Video Streaming with Diffusion EnhancementQihua Zhou, Ruibin Li, Jingcai Guo, Yaodong Huang et al.AAAI 2025 · 2 citations
- Towards High-Quality and Efficient Video Super-Resolution via Spatial-Temporal Data OverfittingGen Li, Jie Ji, Minghai Qin, Wei Niu et al.CVPR 2023
Builds on6
- LAPAR: Linearly-Assembled Pixel-Adaptive Regression Network for Single Image Super-resolution and BeyondWenbo Li, Kun Zhou, Lu Qi, Nianjuan Jiang et al.NeurIPS 2020 · 293 citations
- Learning temporal coherence via self-supervision for GAN-based video generationMengyu Chu, You Xie, Jonas Mayer, Laura Leal-Taixé et al.SIGGRAPH 2020 · 198 citations
- Streaming 360-Degree Videos Using Super-ResolutionMallesham Dasari, Arani Bhattacharya, Santiago Vargas, Pranjal Sahu et al.INFOCOM 2020 · 142 citations
- Neural-Enhanced Live Streaming: Improving Live Video Ingest via Online LearningJaehong Kim, Youngmok Jung, Hyunho Yeo, Juncheol Ye et al.SIGCOMM 2020 · 132 citations
- NEMO: enabling neural-enhanced video streaming on commodity mobile devicesHyunho Yeo, Chan Ju Chong, Youngmok Jung, Juncheol Ye et al.MobiCom 2020 · 118 citations
Related papers
- Efficient Video Compression via Content-Adaptive Super-ResolutionMehrdad Khani Shirkoohi, Vibhaalakshmi Sivaraman, Mohammad AlizadehICCV 2021 · 68 citations
- CASVA: Configuration-Adaptive Streaming for Live Video AnalyticsMiao Zhang, Fangxin Wang, Jiangchuan LiuINFOCOM 2022 · 69 citations
- VidIQ: Inference-Aware Neural Codecs for Quality-Enhanced, Real-Time Video AnalyticsAndong Zhu, Sheng Zhang, Xiaohang Shi, Hesheng Sun et al.ACM MM 2025
- Swift: Adaptive Video Streaming with Layered Neural CodecsMallesham Dasari, Kumara Kahatapitiya, Samir R. Das, Aruna Balasubramanian et al.NSDI 2022
- DeNC++: Efficient Diffusion-Enhanced Neural Codec for End-to-end Semantic Streaming at the EdgeQihua Zhou, Wangjiang Gong, Zili Meng, Yaxiong Xie et al.AAAI 2026
