PartRM: Modeling Part-Level Dynamics with Large Cross-State Reconstruction Model
Mingju Gao, Yike Pan, Huan-ang Gao, Zongzheng Zhang, Wenyi Li, Hao Dong, Hao Tang, Li Yi, Hao Zhao
2025Year
4Top-tier citations
Abstract
Figure 1. We present PartRM, given a single-view image and user-specific drags, PartRM can efficiently models appearance, geometry, and part-level motion in a feed-forward manner. Unlike the previous state-of-the-art, Puppet-Master [24], PartRM achieves higher PSNR and significantly faster inference times. As shown in (b) compared to (a), PartRM produces 3D-aware, part-level motions with enhanced multi-view consistency, delivering more realistic and coherent results across different viewpoints.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- ART: Articulated Reconstruction TransformerZizhang Li, Cheng Zhang, Zhengqin Li, Henry Howard-Jenkins et al.CVPR 2026 · 12 citations
- SPARK: Sim-ready Part-level Articulated Reconstruction with VLM KnowledgeYumeng He, Ying Jiang, Jiayin Lu, Yin Yang et al.CVPR 2026 · 6 citations
- ArtPro: Self-Supervised Articulated Object Reconstruction with Adaptive Integration of Mobility ProposalsXuelu Li, Zhaonan Wang, Xiaogang Wang, Lei Wu et al.CVPR 2026 · 1 citation
- SimArt: Decomposing Monolithic Meshes into Sim-ready Articulated Assets via MLLMChuanrui Zhang, Minghan Qin, Yuang Wang, Baifeng Xie et al.SIGGRAPH 2026
Builds on33
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
Related papers
- Puppet-Master: Scaling Interactive Video Generation as a Motion Prior for Part-Level DynamicsRuining Li, Chuanxia Zheng, Christian Rupprecht, Andrea VedaldiICCV 2025 · 2 citations
- Particulate: Feed-Forward 3D Object ArticulationRuining Li, Yuxin Yao, Chuanxia Zheng, Christian Rupprecht et al.CVPR 2026 · 22 citations
- PARIS: Part-level Reconstruction and Motion Analysis for Articulated ObjectsJiayi Liu, Ali Mahdavi-Amiri, Manolis SavvaICCV 2023 · 103 citations
- HumanRAM: Feed-forward Human Reconstruction and Animation Model using TransformersZhiyuan Yu, Zhe Li, Hujun Bao, Can Yang et al.SIGGRAPH 2025 · 2 citations
- RTGaze: Real-Time 3D-Aware Gaze Redirection from a Single ImageHengfei Wang, Zhongqun Zhang, Yihua Cheng, Hyung Jin ChangAAAI 2026
