To The Point: Correspondence-driven monocular 3D category reconstruction
Filippos Kokkinos, Iasonas Kokkinos
摘要
We present To The Point (TTP), a method for reconstructing 3D objects from a single image using 2D to 3D correspondences learned from weak supervision. We recover a 3D shape from a 2D image by first regressing the 2D positions corresponding to the 3D template vertices and then jointly estimating a rigid camera transform and non-rigid template deformation that optimally explain the 2D positions through the 3D shape projection. By relying on 3D-2D correspondences we use a simple per-sample optimization problem to replace CNN-based regression of camera pose and non-rigid deformation and thereby obtain substantially more accurate 3D reconstructions. We treat this optimization as a differentiable layer and train the whole system in an end-to-end manner. We report systematic quantitative improvements on multiple categories and provide qualitative results comprising diverse shape, pose and texture prediction examples. Project website: https: //fkokkinos.github.io/to_the_point/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- BANMo: Building Animatable 3D Neural Models from Many Casual VideosGengshan Yang, Minh Vo, Natalia Neverova, Deva Ramanan 等CVPR 2022 · 被引用 113 次
- Neural Surface Reconstruction of Dynamic Scenes with Monocular RGB-D CameraHongrui Cai, Wanquan Feng, Xuetao Feng, Yan Wang 等NeurIPS 2022 · 被引用 83 次
- PPR: Physically Plausible Reconstruction from Monocular VideosGengshan Yang, Shuo Yang, John Z. Zhang, Zachary Manchester 等ICCV 2023 · 被引用 41 次
- Do It Yourself: Learning Semantic Correspondence from Pseudo-LabelsOlaf Dünkel, Thomas Wimmer, Christian Theobalt, Christian Rupprecht 等ICCV 2025 · 被引用 4 次
- SAOR: Single-View Articulated Object ReconstructionMehmet Aygün, Oisin Mac AodhaCVPR 2024
它引用的顶会 Paper11
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 被引用 1,139 次
- C3DPO: Canonical 3D Pose Networks for Non-Rigid Structure From MotionDavid Novotný, Nikhila Ravi, Benjamin Graham, Natalia Neverova 等ICCV 2019 · 被引用 126 次
- Canonical Surface Mapping via Geometric Cycle ConsistencyNilesh Kulkarni, Shubham Tulsiani, Abhinav GuptaICCV 2019 · 被引用 104 次
- Deep Non-Rigid Structure From MotionChen Kong, Simon LuceyICCV 2019 · 被引用 72 次
- Online Adaptation for Consistent Mesh Reconstruction in the WildXueting Li, Sifei Liu, Shalini De Mello, Kihwan Kim 等NeurIPS 2020 · 被引用 62 次
相关 Paper
- From Image Collections to Point Clouds With Self-Supervised Shape and Pose NetworksNavaneet K. L., Ansu Mathew, Shashank Kashyap, Wei-Chih Hung 等CVPR 2020
- Topologically-Aware Deformation Fields for Single-View 3D ReconstructionShivam Duggal, Deepak PathakCVPR 2022 · 被引用 30 次
- Canonical 3D Deformer Maps: Unifying parametric and non-parametric methods for dense weakly-supervised category reconstructionDavid Novotný, Roman Shapovalov, Andrea VedaldiNeurIPS 2020 · 被引用 9 次
- ShapeClipper: Scalable 3D Shape Learning from Single-View Images via Geometric and CLIP-Based ConsistencyZixuan Huang, Varun Jampani, Anh Thai, Yuanzhen Li 等CVPR 2023
- Unsupervised Template-assisted Point Cloud Shape Correspondence NetworkJiacheng Deng, Jiahao Lu, Tianzhu ZhangCVPR 2024
