Novel-view Synthesis and Pose Estimation for Hand-Object Interaction from Sparse Views
Wentian Qu, Zhaopeng Cui, Yinda Zhang, Chenyu Meng, Cuixia Ma, Xiaoming Deng, Hongan Wang
摘要
Hand-object interaction understanding and the barely addressed novel view synthesis are highly desired in the immersive communication, whereas it is challenging due to the high deformation of hand and heavy occlusions between hand and object. In this paper, we propose a neural rendering and pose estimation system for hand-object interaction from sparse views, which can also enable 3D hand-object interaction editing. We share the inspiration from recent scene understanding work that shows a scene specific model built beforehand can significantly improve and unblock vision tasks especially when inputs are sparse, and extend it to the dynamic hand-object interaction scenario and propose to solve the problem in two stages. We first learn the shape and appearance prior knowledge of hands and objects separately with the neural representation at the offline stage. During the online stage, we design a rendering-based joint model fitting framework to understand the dynamic hand-object interaction with the pre-built hand and object models as well as interaction priors, which thereby overcomes penetration and separation issues between hand and object and also enables novel view synthesis. In order to get stable contact during the hand-object interaction process in a sequence, we propose a stable contact loss to make the contact region to be consistent. Experiments demonstrate that our method outperforms the state-of-the-art methods. Code and dataset are available in project web-page https://iscas3dv.github.io/HO-NeRF.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- GeneOH Diffusion: Towards Generalizable Hand-Object Interaction Denoising via Denoising DiffusionXueyi Liu, Li YiICLR 2024 · 被引用 35 次
- RoboWheel: A Data Engine from Real-World Human Demonstrations for Cross-Embodiment Robotic LearningYuhong Zhang, Zihan Gao, Shengpeng Li, Ling-Hao Chen 等CVPR 2026 · 被引用 11 次
- OHTA: One-shot Hand Avatar via Data-driven Implicit PriorsXiaozheng Zheng, Chao Wen, Zhuo Su, Zeran Xu 等CVPR 2024 · 被引用 7 次
- ForeHOI: Feed-forward 3D Object Reconstruction from Daily Hand-Object Interaction VideosYuantao Chen, Jiahao Chang, Chongjie Ye, Chaoran Zhang 等CVPR 2026 · 被引用 6 次
- HOGSA: Bimanual Hand-Object Interaction Understanding with 3D Gaussian Splatting Based Data AugmentationWentian Qu, Jiahe Li, Jian Cheng, Jian Shi 等AAAI 2025 · 被引用 4 次
它引用的顶会 Paper37
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman 等ICCV 2021 · 被引用 2,700 次
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt 等NeurIPS 2021 · 被引用 2,500 次
- Nerfies: Deformable Neural Radiance FieldsKeunhong Park, Utkarsh Sinha, Jonathan T. Barron, Sofien Bouaziz 等ICCV 2021 · 被引用 1,442 次
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 被引用 1,421 次
相关 Paper
- Single-view Image to Novel-view Generation for Hand-Object InteractionsZhongqun Zhang, Yihua Cheng, Eduardo Pérez-Pellitero, Yiren Zhou 等AAAI 2025 · 被引用 1 次
- HandNeRF: Neural Radiance Fields for Animatable Interacting HandsZhiyang Guo, Wengang Zhou, Min Wang, Li Li 等CVPR 2023
- Neural Free-Viewpoint Performance Rendering under Complex Human-object InteractionsGuoxing Sun, Xin Chen, Yizhang Chen, Anqi Pang 等ACM MM 2021 · 被引用 37 次
- Hand-Object Interaction Image GenerationHezhen Hu, Weilun Wang, Wengang Zhou, Houqiang LiNeurIPS 2022 · 被引用 24 次
- HOSNeRF: Dynamic Human-Object-Scene Neural Radiance Fields from a Single VideoJia-Wei Liu, Yan-Pei Cao, Tianyuan Yang, Zhongcong Xu 等ICCV 2023 · 被引用 35 次
