HiMoR: Monocular Deformable Gaussian Reconstruction with Hierarchical Motion Representation
Yiming Liang, Tianhan Xu, Yuta Kikuchi
摘要
We present Hierarchical Motion Representation (HiMoR), a novel deformation representation for 3D Gaussian primitives capable of achieving high-quality monocular dynamic 3D reconstruction. The insight behind HiMoR is that motions in everyday scenes can be decomposed into coarser motions that serve as the foundation for finer details. Using a tree structure, HiMoR's nodes represent different levels of motion detail, with shallower nodes modeling coarse motion for temporal smoothness and deeper nodes capturing finer motion. Additionally, our model uses a few shared motion bases to represent motions of different sets of nodes, aligning with the assumption that motion tends to be smooth and simple. This motion representation design provides Gaussians with a more structured deformation, maximizing the use of temporal relationships to tackle the challenging task of monocular dynamic 3D reconstruction. We also propose using a more reliable perceptual metric as an alternative, given that pixel-level metrics for evaluating monocular dynamic 3D reconstruction can sometimes fail to accurately reflect the true quality of reconstruction. Extensive experiments demonstrate our method's efficacy in achieving superior novel view synthesis from challenging monocular videos with complex motions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Orientation-anchored Hyper-Gaussian for 4D Reconstruction from Casual VideosJunyi Wu, Jiachen Tao, Haoxuan Wang, Gaowen Liu 等NeurIPS 2025 · 被引用 9 次
- Uncertainty Matters in Dynamic Gaussian Splatting for Monocular 4D ReconstructionFengzhi Guo, Chih-Chuan Hsu, Sihao Ding, Cheng ZhangICLR 2026 · 被引用 6 次
- From Tokens to Nodes: Semantic-Guided Motion Control for Dynamic 3D Gaussian SplattingJianing Chen, Zehao Li, Yujun Cai, Hao Jiang 等ICLR 2026 · 被引用 3 次
- Dynamic Gaussian Splatting from Defocused and Motion-blurred Monocular VideosXuankai Zhang, Junjin Xiao, Qing ZhangNeurIPS 2025 · 被引用 2 次
- ProDyG: Progressive Dynamic Scene Reconstruction via Gaussian Splatting from Monocular VideosShi Chen, Erik Sandström, Sandro Lombardi, Siyuan Li 等NeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper37
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Nerfies: Deformable Neural Radiance FieldsKeunhong Park, Utkarsh Sinha, Jonathan T. Barron, Sofien Bouaziz 等ICCV 2021 · 被引用 1,442 次
- Depth Anything: Unleashing the Power of Large-Scale Unlabeled DataLihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu 等CVPR 2024 · 被引用 847 次
相关 Paper
- Motion Decoupled 3D Gaussian Splatting for Dynamic Object RepresentationXiao Hu, Libo Long, Jochen LangAAAI 2025 · 被引用 2 次
- HAIF-GS: Hierarchical and Induced Flow-Guided Gaussian Splatting for Dynamic SceneJianing Chen, Zehao Li, Yujun Cai, Hao Jiang 等NeurIPS 2025 · 被引用 14 次
- Kinematics-Driven Gaussian Shape Deformation for Blurry Monocular Dynamic ScenesYeon-Ji Song, Kiyoung Kwon, Junoh Lee, Jin-Hwa Kim 等ICML 2026
- Hi-Gaussian: Hierarchical Gaussians Under Normalized Spherical Projection for Single-View 3D ReconstructionBinjian Xie, Pengju Zhang, Hao Wei, Yihong WuICCV 2025 · 被引用 2 次
- FLAG-4D: Flow-Guided Local-Global Dual-Deformation Model for 4D ReconstructionGuan Yuan Tan, Ngoc Tuan Vu, Arghya Pal, Sailaja Rajanala 等AAAI 2026
