Compact Neural Volumetric Video Representations with Dynamic Codebooks
Haoyu Guo, Sida Peng, Yunzhi Yan, Linzhan Mou, Yujun Shen, Hujun Bao, Xiaowei Zhou
Abstract
This paper addresses the challenge of representing high-fidelity volumetric videos with low storage cost. Some recent feature grid-based methods have shown superior performance of fast learning implicit neural representations from input 2D images. However, such explicit representations easily lead to large model sizes when modeling dynamic scenes. To solve this problem, our key idea is reducing the spatial and temporal redundancy of feature grids, which intrinsically exist due to the self-similarity of scenes. To this end, we propose a novel neural representation, named dynamic codebook, which first merges similar features for the model compression and then compensates for the potential decline in rendering quality by a set of dynamic codes. Experiments on the NHR and DyNeRF datasets demonstrate that the proposed approach achieves state-of-the-art rendering quality, while being able to achieve more storage efficiency. The source code is available at https://github.com/zju3dv/compact_vv .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cdcf2464-b66f-4453-9fe0-7561f5faf3b1Cited by top-tier papers6
- LightGaussian: Unbounded 3D Gaussian Compression with 15x Reduction and 200+ FPSZhiwen Fan, Kevin Wang, Kairun Wen, Zehao Zhu et al.NeurIPS 2024 · 681 citations
- TeTriRF: Temporal Tri-Plane Radiance Fields for Efficient Free-Viewpoint VideoMinye Wu, Zehao Wang, Georgios Kouros, Tinne TuytelaarsCVPR 2024 · 10 citations
- Compressing Streamable Free-Viewpoint Videos to 0.1 MB per FrameLuyang Tang, Jiayu Yang, Rui Peng, Yongqi Zhai et al.AAAI 2025 · 7 citations
- Rate-aware Compression for NeRF-based Volumetric VideoZhiyu Zhang, Guo Lu, Huanxiong Liang, Zhengxue Cheng et al.ACM MM 2024 · 3 citations
- ReMatching Dynamic Reconstruction FlowSara Oblak, Despoina Paschalidou, Sanja Fidler, Matan AtzmonICLR 2025
Builds on27
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman et al.ICCV 2021 · 2,700 citations
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan et al.CVPR 2022 · 1,603 citations
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua et al.NeurIPS 2020 · 1,535 citations
- Plenoxels: Radiance Fields without Neural NetworksSara Fridovich-Keil, Alex Yu, Matthew Tancik, Qinhong Chen et al.CVPR 2022 · 1,237 citations
Related papers
- Neural NeRF CompressionTuan Pham, Stephan MandtICML 2024 · 5 citations
- Learning Neural Volumetric Representations of Dynamic Humans in MinutesChen Geng, Sida Peng, Zhen Xu, Hujun Bao et al.CVPR 2023
- Towards Scalable Neural Representation for Diverse VideosBo He, Xitong Yang, Hanyu Wang, Zuxuan Wu et al.CVPR 2023
- Tree-NeRV: Efficient Non-Uniform Sampling for Neural Video Representation via Tree-Structured Feature GridsJiancheng Zhao, Yifan Zhan, Qingtian Zhu, Mingze Ma et al.ICCV 2025 · 4 citations
- DS-NeRV: Implicit Neural Video Representation with Decomposed Static and Dynamic CodesHao Yan, Zhihui Ke, Xiaobo Zhou, Tie Qiu et al.CVPR 2024 · 18 citations
