Pretrain, Self-train, Distill: A simple recipe for Supersizing 3D Reconstruction
Kalyan Vasudev Alwala, Abhinav Gupta, Shubham Tulsiani
摘要
Our work learns a unified model for single-view 3D reconstruction of objects from hundreds of semantic categories. As a scalable alternative to direct 3D supervision, our work relies on segmented image collections for learning 3D of generic categories. Unlike prior works that use similar supervision but learn independent category-specific models from scratch, our approach of learning a unified model simplifies the training process while also allowing the model to benefit from the common structure across categories. Using image collections from standard recognition datasets, we show that our approach allows learning 3D inference for over 150 object categories. We evaluate using two datasets and qualitatively and quantitatively show that our unified reconstruction approach improves over prior category-specific reconstruction baselines. Our final 3D reconstruction model is also capable of zero-shot inference on images from unseen object categories and we empirically show that increasing the number of training categories improves the reconstruction quality.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Diffusion-SDF: Conditional Generative Modeling of Signed Distance FunctionsGene Chou, Yuval Bahat, Felix HeideICCV 2023 · 被引用 171 次
- Human-3Diffusion: Realistic Avatar Creation via Explicit 3D Consistent Diffusion ModelsYuxuan Xue, Xianghui Xie, Riccardo Marin, Gerard Pons-MollNeurIPS 2024 · 被引用 49 次
- CharacterGen: Efficient 3D Character Generation from Single Images with Multi-View Pose CanonicalizationHao-Yang Peng, Jia-Peng Zhang, Meng-Hao Guo, Yan-Pei Cao 等SIGGRAPH 2024 · 被引用 30 次
- Ross3d: Reconstructive Visual Instruction Tuning With 3D-AwarenessHaochen Wang, Yucheng Zhao, Tiancai Wang, Haoqiang Fan 等ICCV 2025 · 被引用 7 次
- ISS: Image as Stepping Stone for Text-Guided 3D Shape GenerationZhengzhe Liu, Peng Dai, Ruihui Li, Xiaojuan Qi 等ICLR 2023 · 被引用 4 次
它引用的顶会 Paper13
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun 等NeurIPS 2020 · 被引用 1,010 次
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone 等ICCV 2021 · 被引用 686 次
- Escaping Plato's Cave: 3D Shape From Adversarial RenderingPhilipp Henzler, Niloy J. Mitra, Tobias RitschelICCV 2019 · 被引用 254 次
- SDF-SRN: Learning Signed Distance 3D Object Reconstruction from Static ImagesChen-Hsuan Lin, Chaoyang Wang, Simon LuceyNeurIPS 2020 · 被引用 125 次
相关 Paper
- Shelf-Supervised Mesh Prediction in the WildYufei Ye, Shubham Tulsiani, Abhinav GuptaCVPR 2021
- Few-Shot Generalization for Single-Image 3D Reconstruction via PriorsBram Wallace, Bharath HariharanICCV 2019 · 被引用 43 次
- Robust 3D Shape Reconstruction in Zero-Shot from a Single Image in the WildJunhyeong Cho, Kim Youwang, Hunmin Yang, Tae-Hyun OhCVPR 2025
- Multiview Compressive Coding for 3D ReconstructionChao-Yuan Wu, Justin Johnson, Jitendra Malik, Christoph Feichtenhofer 等CVPR 2023
- SAOR: Single-View Articulated Object ReconstructionMehmet Aygün, Oisin Mac AodhaCVPR 2024
