Compressing Streamable Free-Viewpoint Videos to 0.1 MB per Frame
Luyang Tang, Jiayu Yang, Rui Peng, Yongqi Zhai, Shihe Shen, Ronggang Wang
摘要
The success of 3D Gaussian Splatting (3DGS) in static scenes has inspired numerous attempts to construct Free-Viewpoint Videos (FVVs) of dynamic scenes from multi-view videos. Despite advancements in current techniques, simultaneously achieving photo-realistic view synthesis results, fast on-the-fly training, real-time rendering, and low storage costs remains a formidable problem. To address these challenges, we propose the first Gaussian-based streamable FVV intelligent compression framework named iFVC. Specifically, we utilize an anchor-based Gaussian representation to model the scene. To achieve on-the-fly training, we propose a Binary Transformation Cache (BTC) to model the dynamic changes between adjacent timesteps, which not only ensures compactness but also supports precise bit rate estimation. Furthermore, we carefully design a high-resolution transformation tri-plane assisted by a saliency grid as our BTC, allowing for accurate dynamic capture. The entire pipeline is regarded as a joint optimization of rate and distortion to achieve optimal compression performance. Experiments on widely used datasets demonstrate the state-of-the-art performance of our framework in both synthesis quality and efficiency, i.e., achieving per-frame training in 13 seconds with a storage cost of 0.1 MB and real-time rendering at 120 FPS.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Motion Matters: Compact Gaussian Streaming for Free-Viewpoint Video ReconstructionJiacong Chen, Qingyu Mao, Youneng Bao, Xiandong Meng 等NeurIPS 2025 · 被引用 7 次
- SGI: Structured 2D Gaussians for Efficient and Compact Large Image RepresentationZixuan Pan, Kaiyuan Tang, Jun Xia, Yifan Qin 等CVPR 2026 · 被引用 3 次
- ClipGStream: Clip-Stream Gaussian Splatting for Any Length and Any Motion Multi-View Dynamic Scene ReconstructionJie Liang, Jiahao Wu, Chao Wang, Jiayu Yang 等CVPR 2026 · 被引用 3 次
它引用的顶会 Paper27
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Direct Voxel Grid Optimization: Super-fast Convergence for Radiance Fields ReconstructionCheng Sun, Min Sun, Hwann-Tzong ChenCVPR 2022 · 被引用 859 次
- Real-time Photorealistic Dynamic Scene Representation and Rendering with 4D Gaussian SplattingZeyu Yang, Hongye Yang, Zijie Pan, Li ZhangICLR 2024 · 被引用 529 次
- 4D Gaussian Splatting for Real-Time Dynamic Scene RenderingGuanjun Wu, Taoran Yi, Jiemin Fang, Lingxi Xie 等CVPR 2024 · 被引用 513 次
相关 Paper
- StreamSTGS: Streaming Spatial and Temporal Gaussian Grids for Real-Time Free-Viewpoint VideoZhihui Ke, Yuyang Liu, Xiaobo Zhou, Tie QiuAAAI 2026
- 3DGStream: On-the-Fly Training of 3D Gaussians for Efficient Streaming of Photo-Realistic Free-Viewpoint VideosJiakai Sun, Han Jiao, Guangyuan Li, Zhanjie Zhang 等CVPR 2024 · 被引用 72 次
- 4DGC: Rate-Aware 4D Gaussian Compression for Efficient Streamable Free-Viewpoint VideoQiang Hu, Zihan Zheng, Houqiang Zhong, Sihua Fu 等CVPR 2025
- D-FCGS: Feedforward Compression of Dynamic Gaussian Splatting for Free-Viewpoint VideosWenkang Zhang, Yan Zhao, Qiang Wang, Zhixin Xu 等AAAI 2026 · 被引用 1 次
- Fast Feedforward 3D Gaussian Splatting CompressionYihang Chen, Qianyi Wu, Mengyao Li, Weiyao Lin 等ICLR 2025 · 被引用 1 次
