LiFteR: Unleash Learned Codecs in Video Streaming with Loose Frame Referencing
Bo Chen, Zhisheng Yan, Yinjie Zhang, Zhe Yang, Klara Nahrstedt
Abstract
Video codecs are essential for video streaming. While traditional codecs like AVC and HEVC are successful, learned codecs built on deep neural networks (DNNs) are gaining popularity due to their superior coding efficiency and quality of experience (QoE) in video streaming. However, using learned codecs built with sophisticated DNNs in video streaming leads to slow decoding and low frame rate, thereby degrading the QoE. The fundamental problem is the tight frame referencing design adopted by most codecs, which delays the processing of the current frame until its immediate predecessor frame is reconstructed. To overcome this limitation, we propose LiFteR, a novel video streaming system that operates a learned video codec with loose frame referencing (LFR). LFR is a unique frame referencing paradigm that redefines the reference relation between frames and allows parallelism in the learned video codec to boost the frame rate. LiFteR has three key designs: (i) the LFR video dispatcher that routes video data to the codec based on LFR, (ii) LFR learned codec that enhances coding efficiency in LFR with minimal impact on decoding speed, and (iii) streaming supports that enables adaptive bitrate streaming with learned codecs in existing infrastructures. In our evaluation, LiFteR consistently outperforms existing video streaming systems. Compared to the existing best-performing learned and traditional systems, LiFteR demonstrates up to 23.8% and 19.7% QoE gain, respectively. Furthermore, LiFteR achieves up to a 3.2× frame rate improvement through frame rate configuration.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a074f632-665c-4f5a-b9f3-33ce5a379f62Cited by top-tier papers4
- ACE: Sending Burstiness Control for High-Quality Real-time CommunicationXiangjie Huang, Jiayang Xu, Haiping Wang, Hebin Yu et al.SIGCOMM 2025 · 8 citations
- DeNC: Unleash Neural Codecs in Video Streaming with Diffusion EnhancementQihua Zhou, Ruibin Li, Jingcai Guo, Yaodong Huang et al.AAAI 2025 · 2 citations
- SAND: A New Programming Abstraction for Video-based Deep LearningJuncheol Ye, Seungkook Lee, Hwijoon Lim, Jihyuk Lee et al.SOSP 2025
- DeNC++: Efficient Diffusion-Enhanced Neural Codec for End-to-end Semantic Streaming at the EdgeQihua Zhou, Wangjiang Gong, Zili Meng, Yaxiong Xie et al.AAAI 2026
Builds on12
- Is Space-Time Attention All You Need for Video Understanding?Gedas Bertasius, Heng Wang, Lorenzo TorresaniICML 2021 · 2,927 citations
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
- Learning in situ: a randomized experiment in video streamingFrancis Y. Yan, Hudson Ayers, Chenzhi Zhu, Sadjad Fouladi et al.NSDI 2020 · 360 citations
- Server-Driven Video Streaming for Deep Learning InferenceKuntai Du, Ahsan Pervaiz, Xin Yuan, Aakanksha Chowdhery et al.SIGCOMM 2020 · 238 citations
- Video Compression With Rate-Distortion AutoencodersAmirHossein Habibian, Ties van Rozendaal, Jakub M. Tomczak, Taco CohenICCV 2019 · 233 citations
Related papers
- Swift: Adaptive Video Streaming with Layered Neural CodecsMallesham Dasari, Kumara Kahatapitiya, Samir R. Das, Aruna Balasubramanian et al.NSDI 2022
- Reparo: QoE-Aware Live Video Streaming in Low-Rate Networks by Intelligent Frame RecoveryFulin Wang, Qing Li, Wanxin Shi, Gareth Tyson et al.ACM MM 2023 · 10 citations
- ELF-VC: Efficient Learned Flexible-Rate Video CodingOren Rippel, Alexander G. Anderson, Kedar Tatwawadi, Sanjay Nair et al.ICCV 2021 · 137 citations
- Hierarchical B-Frame Video Coding Using Two-Layer CANF Without Motion CodingDavid Alexandre, Hsueh-Ming Hang, Wen-Hsiao PengCVPR 2023
- HyTIP: Hybrid Temporal Information Propagation for Masked Conditional Residual Video CodingYi-Hsin Chen, Yi-Chen Yao, Kuan-Wei Ho, Chun-Hung Wu et al.ICCV 2025 · 3 citations
