HOISDF: Constraining 3D Hand-Object Pose Estimation with Global Signed Distance Fields
Haozhe Qi, Chen Zhao, Mathieu Salzmann, Alexander Mathis
Abstract
Human hands are highly articulated and versatile at handling objects. Jointly estimating the 3D poses of a hand and the object it manipulates from a monocular camera is challenging due to frequent occlusions. Thus, existing methods often rely on intermediate 3D shape representations to increase performance. These representations are typically explicit, such as 3D point clouds or meshes, and thus provide information in the direct surroundings of the intermediate hand pose estimate. To address this, we in-troduce HOISDF, a Signed Distance Field (SDF) guided hand-object pose estimation network, which jointly exploits hand and object SDFs to provide a global, implicit repre-sentation over the complete reconstruction volume. Specif-ically, the role of the SDFs is threefold: equip the visual encoder with implicit shape information, help to encode hand-object interactions, and guide the hand and object pose regression via SDF-based sampling and by augmenting the feature representations. We show that HOISDF achieves state-of-the-art results on hand-object pose esti-mation benchmarks (DexYCB and H03Dv2). Code is avail-able at https://github.com/amathislabIHOISDF.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 664c771e-2f2d-44bc-8b84-ee0d38a6c7e5Cited by top-tier papers6
- Generalizable Hand-Object Modeling from Monocular RGB Images via 3D GaussiansXingyu Liu, Pengfei Ren, Qi Qi, Haifeng Sun et al.NeurIPS 2025 · 5 citations
- HORT: Monocular Hand-held Objects Reconstruction with TransformersZerui Chen, Rolandos Alexandros Potamias, Shizhe Chen, Cordelia SchmidICCV 2025 · 4 citations
- Towards Dynamic 3D Reconstruction of Hand-Instrument Interaction in Ophthalmic SurgeryMing Hu, Zhengdi Yu, Feilong Tang, Kaiwen Chen et al.NeurIPS 2025 · 1 citation
- Diffusion-Based 3D Hand Motion Recovery with Intuitive PhysicsYufei Zhang, Zijun Cui, Jeffrey O. Kephart, Qiang JiICCV 2025 · 1 citation
- Clay-to-Stone: Phase-wise 3D Gaussian Splatting for Monocular Articulated Hand-Object Manipulation ModelingXingyu Liu, Pengfei Ren, Qi Qi, Haifeng Sun et al.CVPR 2026 · 1 citation
Builds on25
- Masked Autoencoders As Spatiotemporal LearnersChristoph Feichtenhofer, Haoqi Fan, Yanghao Li, Kaiming HeNeurIPS 2022 · 690 citations
- FreiHAND: A Dataset for Markerless Capture of Hand Pose and Shape From Single RGB ImagesChristian Zimmermann, Duygu Ceylan, Jimei Yang, Bryan C. Russell et al.ICCV 2019 · 493 citations
- Neural Unsigned Distance Fields for Implicit Function LearningJulian Chibane, Aymen Mir, Gerard Pons-MollNeurIPS 2020 · 415 citations
- StyleSDF: High-Resolution 3D-Consistent Image and Geometry GenerationRoy Or-El, Xuan Luo, Mengyi Shan, Eli Shechtman et al.CVPR 2022 · 229 citations
- A2J: Anchor-to-Joint Regression Network for 3D Articulated Pose Estimation From a Single Depth ImageFu Xiong, Boshen Zhang, Yang Xiao, Zhiguo Cao et al.ICCV 2019 · 178 citations
Related papers
- gSDF: Geometry-Driven Signed Distance Functions for 3D Hand-Object ReconstructionZerui Chen, Shizhe Chen, Cordelia Schmid, Ivan LaptevCVPR 2023
- DDF-HO: Hand-Held Object Reconstruction via Conditional Directed Distance FieldChenyangguang Zhang, Yan Di, Ruida Zhang, Guangyao Zhai et al.NeurIPS 2023 · 22 citations
- Tracking and Reconstructing Hand Object Interactions from Point Cloud Sequences in the WildJiayi Chen, Mi Yan, Jiazhao Zhang, Yinzhen Xu et al.AAAI 2023 · 26 citations
- Learning Explicit Contact for Implicit Reconstruction of Hand-Held Objects from Monocular ImagesJunxing Hu, Hongwen Zhang, Zerui Chen, Mengcheng Li et al.AAAI 2024 · 15 citations
- PHRIT: Parametric Hand Representation with Implicit TemplateZhisheng Huang, Yujin Chen, Di Kang, Jinlu Zhang et al.ICCV 2023 · 9 citations
