STaR: Self-Supervised Tracking and Reconstruction of Rigid Objects in Motion With Neural Rendering
Wentao Yuan, Zhaoyang Lv, Tanner Schmidt, Steven Lovegrove
摘要
We present STaR, a novel method that performs Self-supervised Tracking and Reconstruction of dynamic scenes with rigid motion from multi-view RGB videos without any manual annotation. Recent work has shown that neural networks are surprisingly effective at the task of compressing many views of a scene into a learned function which maps from a viewing ray to an observed radiance value via volume rendering. Unfortunately, these methods lose all their predictive power once any object in the scene has moved. In this work, we explicitly model rigid motion of objects in the context of neural representations of radiance fields. We show that without any additional human specified supervision, we can reconstruct a dynamic scene with a single rigid object in motion by simultaneously decomposing it into its two constituent parts and encoding each with its own neural representation. We achieve this by jointly optimizing the parameters of two neural radiance fields and a set of rigid poses which align the two fields at each frame. On both synthetic and real world datasets, we demonstrate that our method can render photorealistic novel views, where novelty is measured on both spatial and temporal axes. Our factored representation furthermore enables animation of unseen object motion.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper38
- Neural 3D Video Synthesis from Multi-view VideoTianye Li, Mira Slavcheva, Michael Zollhöfer, Simon Green 等CVPR 2022 · 被引用 324 次
- NeRFPlayer: A Streamable Dynamic Scene Representation with Decomposed Neural Radiance FieldsLiangchen Song, Anpei Chen, Zhong Li, Zhang Chen 等IEEE VR 2023 · 被引用 246 次
- Neural Articulated Radiance FieldAtsuhiro Noguchi, Xiao Sun, Stephen Lin, Tatsuya HaradaICCV 2021 · 被引用 242 次
- Neural Human Performer: Learning Generalizable Radiance Fields for Human Performance RenderingYoungjoong Kwon, Dahun Kim, Duygu Ceylan, Henry FuchsNeurIPS 2021 · 被引用 224 次
- D^2NeRF: Self-Supervised Decoupling of Dynamic and Static Objects from a Monocular VideoTianhao Wu, Fangcheng Zhong, Andrea Tagliasacchi, Forrester Cole 等NeurIPS 2022 · 被引用 184 次
它引用的顶会 Paper13
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 被引用 2,416 次
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua 等NeurIPS 2020 · 被引用 1,535 次
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun 等NeurIPS 2020 · 被引用 1,010 次
相关 Paper
- Non-Rigid Neural Radiance Fields: Reconstruction and Novel View Synthesis of a Dynamic Scene From Monocular VideoEdgar Tretschk, Ayush Tewari, Vladislav Golyanik, Michael Zollhöfer 等ICCV 2021 · 被引用 617 次
- MonoNeRF: Learning Generalizable NeRFs from Monocular Videos without Camera PosesYang Fu, Ishan Misra, Xiaolong WangICML 2023 · 被引用 13 次
- Neural Scene Graphs for Dynamic ScenesJulian Ost, Fahim Mannan, Nils Thuerey, Julian Knodt 等CVPR 2021
- Neural Reconstruction of Relightable Human Model from Monocular VideoWenzhang Sun, Yunlong Che, Yandong Guo, Han HuangICCV 2023 · 被引用 20 次
- Space-Time Neural Irradiance Fields for Free-Viewpoint VideoWenqi Xian, Jia-Bin Huang, Johannes Kopf, Changil KimCVPR 2021
