Learning Compositional Representation for 4D Captures With Neural ODE
Boyan Jiang, Yinda Zhang, Xingkui Wei, Xiangyang Xue, Yanwei Fu
摘要
Learning based representation has become the key to the success of many computer vision systems. While many 3D representations have been proposed, it is still an unaddressed problem how to represent a dynamically changing 3D object. In this paper, we introduce a compositional representation for 4D captures, i.e. a deforming 3D object over a temporal span, that disentangles shape, initial state, and motion respectively. Each component is represented by a latent code via a trained encoder. To model the motion, a neural Ordinary Differential Equation (ODE) is trained to update the initial state conditioned on the learned motion code, and a decoder takes the shape code and the updated state code to reconstruct the 3D model at each time stamp. To this end, we propose an Identity Exchange Training (IET) strategy to encourage the network to learn effectively decoupling each component. Extensive experiments demonstrate that the proposed method outperforms existing stateof-the-art deep learning based methods on 4D reconstruction, and significantly improves on various tasks, including motion transfer and completion.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Neural Temporal Walks: Motif-Aware Representation Learning on Continuous-Time Dynamic GraphsMing Jin, Yuan-Fang Li, Shirui PanNeurIPS 2022 · 被引用 130 次
- NeMF: Neural Motion Fields for Kinematic AnimationChengan He, Jun Saito, James Zachary, Holly E. Rushmeier 等NeurIPS 2022 · 被引用 84 次
- CaDeX: Learning Canonical Deformation Coordinate Space for Dynamic Surface Representation via Neural HomeomorphismJiahui Lei, Kostas DaniilidisCVPR 2022 · 被引用 37 次
- No Pain, Big Gain: Classify Dynamic Point Cloud Sequences with Static Models by Fitting Feature-level Space-time SurfacesJia-Xing Zhong, Kaichen Zhou, Qingyong Hu, Bing Wang 等CVPR 2022 · 被引用 22 次
- H4D: Human 4D Modeling by Learning Neural Compositional RepresentationBoyan Jiang, Yinda Zhang, Xingkui Wei, Xiangyang Xue 等CVPR 2022 · 被引用 20 次
它引用的顶会 Paper9
- Swapping Autoencoder for Deep Image ManipulationTaesung Park, Jun-Yan Zhu, Oliver Wang, Jingwan Lu 等NeurIPS 2020 · 被引用 376 次
- Occupancy Flow: 4D Reconstruction by Learning Particle DynamicsMichael Niemeyer, Lars M. Mescheder, Michael Oechsle, Andreas GeigerICCV 2019 · 被引用 314 次
- Pixel2Mesh++: Multi-View 3D Mesh Generation via DeformationChao Wen, Yinda Zhang, Zhuwen Li, Yanwei FuICCV 2019 · 被引用 279 次
- XNect: real-time multi-person 3D motion capture with a single RGB cameraDushyant Mehta, Oleksandr Sotnychenko, Franziska Mueller, Weipeng Xu 等SIGGRAPH 2020 · 被引用 267 次
- Learning Compositional Representations for Few-Shot RecognitionPavel Tokmakov, Yu-Xiong Wang, Martial HebertICCV 2019 · 被引用 133 次
相关 Paper
- The Structure-Equivalent Prior: Unifying Temporal Dynamics and 3D Evolution in 4D Latent SpaceJingyuan Gao, Tianyu Shen, Ruosen Hao, Te Guo 等AAAI 2026
- Velox: Learning Representations of 4D Geometry and AppearanceAnagh Malik, Dorian Chan, Xiaoming Zhao, David B. Lindell 等CVPR 2026
- Mesh4D: 4D Mesh Reconstruction and Tracking from Monocular VideoZeren Jiang, Chuanxia Zheng, Iro Laina, Diane Larlus 等CVPR 2026 · 被引用 13 次
- Inferring Compositional 4D Scenes without Ever Seeing OneAhmet Berke Gökmen, Ajad Chhatkuli, Luc Van Gool, Danda PaudelCVPR 2026 · 被引用 1 次
- Learning Parallel Dense Correspondence From Spatio-Temporal Descriptors for Efficient and Robust 4D ReconstructionJiapeng Tang, Dan Xu, Kui Jia, Lei ZhangCVPR 2021
