VRVVC: Variable-Rate NeRF-Based Volumetric Video Compression
Qiang Hu, Houqiang Zhong, Zihan Zheng, Xiaoyun Zhang, Zhengxue Cheng, Li Song, Guangtao Zhai, Yanfeng Wang
Abstract
Neural Radiance Field (NeRF)-based volumetric video has revolutionized visual media by delivering photorealistic Free-Viewpoint Video (FVV) experiences that provide audiences with unprecedented immersion and interactivity. However, the substantial data volumes pose significant challenges for storage and transmission. Existing solutions typically optimize NeRF representation and compression independently or focus on a single fixed rate-distortion (RD) tradeoff. In this paper, we propose VRVVC, a novel end-to-end joint optimization variable-rate framework for volumetric video compression that achieves variable bitrates using a single model while maintaining superior RD performance. Specifically, VRVVC introduces a compact tri-plane implicit residual representation for inter-frame modeling of long-duration dynamic scenes, effectively reducing temporal redundancy. We further propose a variable-rate residual representation compression scheme that leverages a learnable quantization and a tiny MLP-based entropy model. This approach enables variable bitrates through the utilization of predefined Lagrange multipliers to manage the quantization error of all latent representations. Finally, we present an end-to-end progressive training strategy combined with a multi-rate-distortion loss function to optimize the entire framework. Extensive experiments demonstrate that VRVVC achieves a wide range of variable bitrates within a single model and surpasses the RD performance of existing methods across various datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext df16699c-9a6a-473e-82e9-d3e2703f1935Cited by top-tier papers4
- 4DGCPro: Efficient Hierarchical 4D Gaussian Compression for Progressive Volumetric Video StreamingZihan Zheng, Zhenlong Wu, Houqiang Zhong, Yuan Tian et al.NeurIPS 2025 · 12 citations
- Semantics Versus Identity: A Divide-and-Conquer Approach Towards Adjustable Medical Image De-IdentificationYuan Tian, Shuo Wang, Rongzhao Zhang, Zijian Chen et al.ICCV 2025 · 3 citations
- BEAM: Bridging Physically-based Rendering and Gaussian Modeling for Relightable Volumetric VideoYu Hong, Yize Wu, Zhehao Shen, Chengcheng Guo et al.ACM MM 2025 · 1 citation
- StreamSTGS: Streaming Spatial and Temporal Gaussian Grids for Real-Time Free-Viewpoint VideoZhihui Ke, Yuyang Liu, Xiaobo Zhou, Tie QiuAAAI 2026
Builds on24
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Nerfies: Deformable Neural Radiance FieldsKeunhong Park, Utkarsh Sinha, Jonathan T. Barron, Sofien Bouaziz et al.ICCV 2021 · 1,442 citations
- Neural Radiance Flow for 4D View Synthesis and Video ProcessingYilun Du, Yinan Zhang, Hong-Xing Yu, Joshua B. Tenenbaum et al.ICCV 2021 · 329 citations
- Variable Rate Deep Image Compression With a Conditional AutoencoderYoojin Choi, Mostafa El-Khamy, Jungwon LeeICCV 2019 · 265 citations
Related papers
- HPC: Hierarchical Progressive Coding Framework for Volumetric VideoZihan Zheng, Houqiang Zhong, Qiang Hu, Xiaoyun Zhang et al.ACM MM 2024 · 9 citations
- Rate-aware Compression for NeRF-based Volumetric VideoZhiyu Zhang, Guo Lu, Huanxiong Liang, Zhengxue Cheng et al.ACM MM 2024 · 3 citations
- TeTriRF: Temporal Tri-Plane Radiance Fields for Efficient Free-Viewpoint VideoMinye Wu, Zehao Wang, Georgios Kouros, Tinne TuytelaarsCVPR 2024 · 10 citations
- Neural Residual Radiance Fields for Streamably Free-Viewpoint VideosLiao Wang, Qiang Hu, Qihan He, Ziyu Wang et al.CVPR 2023
- NeRFCodec: Neural Feature Compression Meets Neural Radiance Fields for Memory-Efficient Scene RepresentationSicheng Li, Hao Li, Yiyi Liao, Lu YuCVPR 2024
