3D-LFM: Lifting Foundation Model
Mosam Dabhi, László A. Jeni, Simon Lucey
Abstract
Abstract The lifting of a 3D structure and camera from 2D landmarks is at the cornerstone of the discipline of computer vision. Traditional methods have been confined to specific rigid objects, such as those in Perspective-n-Point (PnP) problems, but deep learning has expanded our capability to reconstruct a wide range of object classes (e.g. C3DPO [18] and PAUL [24] ) with resilience to noise, occlusions, and perspective distortions. However, all these techniques have been limited by the fundamental need to establish correspondences across the 3D training data, significantly limiting their utility to applications where one has an abundance of "in-correspondence" 3D data. Our approach harnesses the inherent permutation equivariance of transformers to manage varying numbers of points per 3D data instance, withstands occlusions, and generalizes
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- 2D-LFM: Lifting Foundation Model without 3D SupervisionMosam Dabhi, Irhas Gill, László A. Jeni, Simon LuceyCVPR 2026
- AniMer: Animal Pose and Shape Estimation Using Family Aware TransformerJin Lyu, Tianyi Zhu, Yi Gu, Li Lin et al.CVPR 2025
Builds on7
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll et al.ICCV 2019 · 1,784 citations
- Human Motion Diffusion ModelGuy Tevet, Sigal Raab, Brian Gordon, Yonatan Shafir et al.ICLR 2023 · 167 citations
- C3DPO: Canonical 3D Pose Networks for Non-Rigid Structure From MotionDavid Novotný, Nikhila Ravi, Benjamin Graham, Natalia Neverova et al.ICCV 2019 · 126 citations
- Deep Non-Rigid Structure From MotionChen Kong, Simon LuceyICCV 2019 · 72 citations
- Animal3D: A Comprehensive Dataset of 3D Animal Pose and ShapeJiacong Xu, Yi Zhang, Jiawei Peng, Wufei Ma et al.ICCV 2023 · 55 citations
Related papers
- PAUL: Procrustean Autoencoder for Unsupervised LiftingChaoyang Wang, Simon LuceyCVPR 2021
- PCLs: Geometry-Aware Neural Reconstruction of 3D Pose With Perspective Crop LayersFrank Yu, Mathieu Salzmann, Pascal Fua, Helge RhodinCVPR 2021
- Multi-View Representation is What You Need for Point-Cloud Pre-TrainingSiming Yan, Chen Song, Youkang Kong, Qixing HuangICLR 2024 · 6 citations
- PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape PredictionPeng Wang, Hao Tan, Sai Bi, Yinghao Xu et al.ICLR 2024 · 170 citations
- Learning an Effective Equivariant 3D Descriptor Without SupervisionRiccardo Spezialetti, Samuele Salti, Luigi Di StefanoICCV 2019 · 41 citations
