GauMVC: Generative Decoupled Gaussian Representation for Human-centric Multi-view Video Compression
Ruoke Yan, Mingjia Yang, Xinfeng Zhang, Haocheng Tang, Qian Yin, Zhipin Deng, Kai Zhang, Li zhang, Siwei Ma
Abstract
Human-centric multi-view video has a clear semantic structure: a static background and dynamic human motion. We propose a generative compression framework that explicitly decouples these components. The background is modeled once with 3D Gaussian Splatting, while the human is represented by a personalized Gaussian avatar reconstructed from a sparse set of key views that are transmitted only once and driven by compact per-frame pose parameters from the Skinned Multi-Person Linear (SMPL) model. The encoder sends only three elements: the background, the key views, and the SMPL parameters, enabling high-fidelity multi-viewpoint synthesis at dramatically reduced bitrates. This shifts compression from low-level redundancy removal to semantics-aware generative modeling. Experiments across multiple human-centric datasets demonstrate superior rate–distortion performance, particularly for long and densely captured sequences, and naturally enable semantic editing.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on13
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- 4D Gaussian Splatting for Real-Time Dynamic Scene RenderingGuanjun Wu, Taoran Yi, Jiemin Fang, Lingxi Xie et al.CVPR 2024 · 513 citations
- PyMAF: 3D Human Pose and Shape Regression with Pyramidal Mesh Alignment Feedback LoopHongwen Zhang, Yating Tian, Xinchi Zhou, Wanli Ouyang et al.ICCV 2021 · 376 citations
- Deformable 3D Gaussians for High-Fidelity Monocular Dynamic Scene ReconstructionZiyi Yang, Xinyu Gao, Wen Zhou, Shaohui Jiao et al.CVPR 2024 · 302 citations
- Compressed 3D Gaussian Splatting for Accelerated Novel View SynthesisSimon Niedermayr, Josef Stumpfegger, Rüdiger WestermannCVPR 2024 · 138 citations
Related papers
- HGC-Avatar: Hierarchical Gaussian Compression for Streamable Dynamic 3D AvatarsHaocheng Tang, Ruoke Yan, Xinhui Yin, Qi Zhang et al.ACM MM 2025 · 3 citations
- High-Fidelity Mobile Avatars with Pruned Local BlendshapesYouyi Zhan, He Wang, Tianjia Shao, Kun ZhouCVPR 2026 · 3 citations
- GART: Gaussian Articulated Template ModelsJiahui Lei, Yufu Wang, Georgios Pavlakos, Lingjie Liu et al.CVPR 2024
- D-FCGS: Feedforward Compression of Dynamic Gaussian Splatting for Free-Viewpoint VideosWenkang Zhang, Yan Zhao, Qiang Wang, Zhixin Xu et al.AAAI 2026 · 1 citation
- Learning Efficient and Generalizable Human Representation with Human Gaussian ModelYifan Liu, Shengjun Zhang, Chensheng Dai, Yang Chen et al.ICCV 2025 · 1 citation
