GauMVC: Generative Decoupled Gaussian Representation for Human-centric Multi-view Video Compression
Ruoke Yan, Mingjia Yang, Xinfeng Zhang, Haocheng Tang, Qian Yin, Zhipin Deng, Kai Zhang, Li zhang, Siwei Ma
摘要
Human-centric multi-view video has a clear semantic structure: a static background and dynamic human motion. We propose a generative compression framework that explicitly decouples these components. The background is modeled once with 3D Gaussian Splatting, while the human is represented by a personalized Gaussian avatar reconstructed from a sparse set of key views that are transmitted only once and driven by compact per-frame pose parameters from the Skinned Multi-Person Linear (SMPL) model. The encoder sends only three elements: the background, the key views, and the SMPL parameters, enabling high-fidelity multi-viewpoint synthesis at dramatically reduced bitrates. This shifts compression from low-level redundancy removal to semantics-aware generative modeling. Experiments across multiple human-centric datasets demonstrate superior rate–distortion performance, particularly for long and densely captured sequences, and naturally enable semantic editing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- 4D Gaussian Splatting for Real-Time Dynamic Scene RenderingGuanjun Wu, Taoran Yi, Jiemin Fang, Lingxi Xie 等CVPR 2024 · 被引用 513 次
- PyMAF: 3D Human Pose and Shape Regression with Pyramidal Mesh Alignment Feedback LoopHongwen Zhang, Yating Tian, Xinchi Zhou, Wanli Ouyang 等ICCV 2021 · 被引用 376 次
- Deformable 3D Gaussians for High-Fidelity Monocular Dynamic Scene ReconstructionZiyi Yang, Xinyu Gao, Wen Zhou, Shaohui Jiao 等CVPR 2024 · 被引用 302 次
- Compressed 3D Gaussian Splatting for Accelerated Novel View SynthesisSimon Niedermayr, Josef Stumpfegger, Rüdiger WestermannCVPR 2024 · 被引用 138 次
相关 Paper
- HGC-Avatar: Hierarchical Gaussian Compression for Streamable Dynamic 3D AvatarsHaocheng Tang, Ruoke Yan, Xinhui Yin, Qi Zhang 等ACM MM 2025 · 被引用 3 次
- High-Fidelity Mobile Avatars with Pruned Local BlendshapesYouyi Zhan, He Wang, Tianjia Shao, Kun ZhouCVPR 2026 · 被引用 3 次
- GART: Gaussian Articulated Template ModelsJiahui Lei, Yufu Wang, Georgios Pavlakos, Lingjie Liu 等CVPR 2024
- D-FCGS: Feedforward Compression of Dynamic Gaussian Splatting for Free-Viewpoint VideosWenkang Zhang, Yan Zhao, Qiang Wang, Zhixin Xu 等AAAI 2026 · 被引用 1 次
- Learning Efficient and Generalizable Human Representation with Human Gaussian ModelYifan Liu, Shengjun Zhang, Chensheng Dai, Yang Chen 等ICCV 2025 · 被引用 1 次
