Common Pets in 3D: Dynamic New-View Synthesis of Real-Life Deformable Categories
Samarth Sinha, Roman Shapovalov, Jeremy Reizenstein, Ignacio Rocco, Natalia Neverova, Andrea Vedaldi, David Novotný
Abstract
Reconstruct unseen videos at test-time Train a category-level model on videos of non-rigid objects TrackeRF TrackeRF Figure 1. We tackle the problem of synthesising new views of deformable objects given only a small number of views taken at different times. We introduce new benchmark data for this task: Common Pets in 3D (CoP3D), containing 4,200 smartphone videos of cats and dogs collected 'in the wild'. We also propose a new method, Tracker-NeRF, a deformable new-view synthesis algorithm which learns a category-level reconstruction prior from videos and applies it to reconstruct new objects at test time.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c116eecd-a7c7-4f4b-bfdb-c6e809c77964Cited by top-tier papers11
- 4DGT: Learning a 4D Gaussian Transformer Using Real-World Monocular VideosZhen Xu, Zhengqin Li, Zhao Dong, Xiaowei Zhou et al.NeurIPS 2025 · 51 citations
- NeoVerse: Enhancing 4D World Model with in-the-wild Monocular VideosYuxue Yang, Lue Fan, Ziqi Shi, Junran Peng et al.CVPR 2026 · 42 citations
- Fast Encoder-Based 3D from Casual Videos via Point Track ProcessingYoni Kasten, Wuyue Lu, Haggai MaronNeurIPS 2024 · 16 citations
- DynamicVerse: A Physically-Aware Multimodal Framework for 4D World ModelingKairun Wen, Yuzhi Huang, Runyu Chen, Hui Zheng et al.NeurIPS 2025 · 11 citations
- SpatialTrackerV2: Advancing 3D Point Tracking with Explicit Camera MotionYuxi Xiao, Jianyuan Wang, Nan Xue, Nikita Karaev et al.ICCV 2025 · 6 citations
Builds on30
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua et al.NeurIPS 2020 · 1,535 citations
- Plenoxels: Radiance Fields without Neural NetworksSara Fridovich-Keil, Alex Yu, Matthew Tancik, Qinhong Chen et al.CVPR 2022 · 1,237 citations
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun et al.NeurIPS 2020 · 1,010 citations
- Block-NeRF: Scalable Large Scene Neural View SynthesisMatthew Tancik, Vincent Casser, Xinchen Yan, Sabeek Pradhan et al.CVPR 2022 · 702 citations
Related papers
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone et al.ICCV 2021 · 686 citations
- Total-Recon: Deformable Scene Reconstruction for Embodied View SynthesisChonghyuk Song, Gengshan Yang, Kangle Deng, Jun-Yan Zhu et al.ICCV 2023 · 27 citations
- D-NeRF: Neural Radiance Fields for Dynamic ScenesAlbert Pumarola, Enric Corona, Gerard Pons-Moll, Francesc Moreno-NoguerCVPR 2021
- Common3D: Self-Supervised Learning of 3D Morphable Models for Common Objects in Neural Feature SpaceLeonhard Sommer, Olaf Dünkel, Christian Theobalt, Adam KortylewskiCVPR 2025
- CodeNeRF: Disentangled Neural Radiance Fields for Object CategoriesWonbong Jang, Lourdes AgapitoICCV 2021 · 246 citations
