ShapeClipper: Scalable 3D Shape Learning from Single-View Images via Geometric and CLIP-Based Consistency
Zixuan Huang, Varun Jampani, Anh Thai, Yuanzhen Li, Stefan Stojanov, James M. Rehg
2023Year
4Top-tier citations
Abstract
Google Research Figure 1. We propose a method that reconstructs 3D object shape from single view real-world images. Our method learns high-quality reconstruction through single-view supervision without known viewpoint and can reconstruct shapes of various objects.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d6fc0aa0-5d86-4914-a8d4-5c4521531003Cited by top-tier papers4
- Primitive-Based 3D Human-Object Interaction Modelling and ProgrammingSiqi Liu, Yong-Lu Li, Zhou Fang, Xinpeng Liu et al.AAAI 2024 · 8 citations
- DSO: Aligning 3D Generators with Simulation Feedback for Physical SoundnessRuining Li, Chuanxia Zheng, Christian Rupprecht, Andrea VedaldiICCV 2025 · 2 citations
- Robust 3D Shape Reconstruction in Zero-Shot from a Single Image in the WildJunhyeong Cho, Kim Youwang, Hunmin Yang, Tae-Hyun OhCVPR 2025
- PointInfinity: Resolution-Invariant Point Diffusion ModelsZixuan Huang, Justin Johnson, Shoubhik Debnath, James M. Rehg et al.CVPR 2024
Builds on20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 1,421 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
Related papers
- From Image Collections to Point Clouds With Self-Supervised Shape and Pose NetworksNavaneet K. L., Ansu Mathew, Shashank Kashyap, Wei-Chih Hung et al.CVPR 2020
- Shape-Pose Ambiguity in Learning 3D Reconstruction from ImagesYunjie Wu, Zhengxing Sun, Youcheng Song, Yunhan Sun et al.AAAI 2021 · 2 citations
- Single Image Shape-from-SilhouettesYawen Lu, Yuxing Wang, Guoyu LuACM MM 2020 · 6 citations
- Few-Shot Generalization for Single-Image 3D Reconstruction via PriorsBram Wallace, Bharath HariharanICCV 2019 · 43 citations
- Pretrain, Self-train, Distill: A simple recipe for Supersizing 3D ReconstructionKalyan Vasudev Alwala, Abhinav Gupta, Shubham TulsianiCVPR 2022 · 23 citations
