PostureHMR: Posture Transformation for 3D Human Mesh Recovery
Yu-Pei Song, Xiao Wu, Zhaoquan Yuanl, Jian-Jun Qiao, Qiang Peng
Abstract
Human Mesh Recovery (HMR) aims to estimate the 3D human body from 2D images, which is a challenging task due to inherent ambiguities in translating 2D observations to 3D space. A novel approach called PostureHMR is pro-posed to leverage a multi-step diffusion-style process, which converts this task into a posture transformation from an SMPL T-pose mesh to the target mesh. To inject the learning process of posture transformation with the physical structure of the human body model, a kinematics-based forward process is proposed to interpolate the intermediate state with pose and shape decomposition. Moreover, a mesh-to-posture (M2P) decoder is designed, by combining the in-put of 3D and 2D mesh constraints estimated from the im-age to model the posture changes in the reverse process. It mitigates the difficulties of posture change learning directly from RGB pixels. To overcome the limitation of pixel-level misalignment of modeling results with the input image, a new trimap-based rendering loss is designed to highlight the areas with poor recognition. Experiments conducted on three widely used datasets demonstrate that the proposed approach outperforms the state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3dc9a6fc-4ce4-4ea5-99b5-839a58a449beCited by top-tier papers4
- PS-Mamba: Spatial-Temporal Graph Mamba for Pose Sequence RefinementHaoye Dong, Gim Hee LeeICCV 2025
- HeatFormer: A Neural Optimizer for Multiview Human Mesh RecoveryYuto Matsubara, Ko NishinoCVPR 2025
- PI-HMR: Towards Robust In-bed Temporal Human Shape Reconstruction with Contact Pressure SensingZiyu Wu, Yufan Xiong, Mengting Niu, Fangting Xie et al.CVPR 2025
- CLEP: Contrastive Language-Pose PretrainingSen Jia, Huayu Wang, Hsiang-Wei Huang, Zhaochong An et al.CVPR 2026
Builds on29
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 1,139 citations
- Label-Efficient Semantic Segmentation with Diffusion ModelsDmitry Baranchuk, Andrey Voynov, Ivan Rubachev, Valentin Khrulkov et al.ICLR 2022 · 700 citations
Related papers
- Skeleton2Mesh: Kinematics Prior Injected Unsupervised Human Mesh RecoveryZhenbo Yu, Junjie Wang, Jingwei Xu, Bingbing Ni et al.ICCV 2021 · 27 citations
- Distribution-Aligned Diffusion for Human Mesh RecoveryLin Geng Foo, Jia Gong, Hossein Rahmani, Jun LiuICCV 2023 · 37 citations
- GenHMR: Generative Human Mesh RecoveryMuhammad Usama Saleem, Ekkasit Pinyoanuntapong, Pu Wang, Hongfei Xue et al.AAAI 2025 · 8 citations
- Score-Guided Diffusion for 3D Human RecoveryAnastasis Stathopoulos, Ligong Han, Dimitris N. MetaxasCVPR 2024 · 15 citations
- Tracking People with 3D RepresentationsJathushan Rajasegaran, Georgios Pavlakos, Angjoo Kanazawa, Jitendra MalikNeurIPS 2021 · 30 citations
