Democratising 2D Sketch to 3D Shape Retrieval Through Pivoting
Pinaki Nath Chowdhury, Ayan Kumar Bhunia, Aneeshan Sain, Subhadeep Koley, Tao Xiang, Yi-Zhe Song
摘要
This paper studies the problem of 2D sketch to 3D shape retrieval, but with a focus on democratising the process. We would like this democratisation to happen on two fronts: (i) to remove the need for large-scale specifically sourced 2D sketch and 3D shape datasets, and (ii) to remove restrictions on how well the user needs to sketch and from what viewpoints. The end result is a system that is trainable using existing datasets, and once trained allows users to sketch regardless of drawing skills and without restriction on view angle. We achieve all this via a clever use of pivoting, along with novel designs that injects 3D understanding of 2D sketches into the system. We perform pivoting using two existing datasets, each from a distant research domain to the other: 2D sketch and photo pairs from the sketch-based image retrieval field (SBIR), and 3D shapes from ShapeNet. It follows that the actual feature pivoting happens on photos from the former and 2D projections from the latter. Doing this already achieves most of our democratisation challenge – the level of 2D sketch abstraction embedded in SBIR dataset offers demoralization on drawing quality, and the whole thing works without a specifically sourced 2D sketch and 3D model pair. To further achieve democratisation on sketching viewpoint, we “lift” 2D sketches to 3D space using Blind Perspective-n-Points (BPnP) that injects 3D-aware information into the sketch encoder. Results show ours achieves competitive performance compared with fully-supervised baselines, while meeting all set democratisation goals.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- AirSketch: Generative Motion to SketchHui Xian Grace Lim, Xuanming Cui, Yogesh S. Rawat, Ser Nam LimNeurIPS 2024 · 被引用 4 次
- Doodle Your 3D: from Abstract Freehand Sketches to Precise 3D ShapesHmrishav Bandyopadhyay, Subhadeep Koley, Ayan Das, Ayan Kumar Bhunia 等CVPR 2024
- Doodle Your Keypoints: Sketch-Based Few-Shot Keypoint DetectionSubhajit Maity, Ayan Kumar Bhunia, Subhadeep Koley, Pinaki Nath Chowdhury 等ICCV 2025
- Sketch Down the FLOPs: Towards Efficient Networks for Human SketchAneeshan Sain, Subhajit Maity, Pinaki Nath Chowdhury, Subhadeep Koley 等CVPR 2025
它引用的顶会 Paper18
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- CDPN: Coordinates-Based Disentangled Pose Network for Real-Time RGB-Based 6-DoF Object Pose EstimationZhigang Li, Gu Wang, Xiangyang JiICCV 2019 · 被引用 482 次
- EPro-PnP: Generalized End-to-End Probabilistic Perspective-n-Points for Monocular Object Pose EstimationHansheng Chen, Pichao Wang, Fan Wang, Wei Tian 等CVPR 2022 · 被引用 175 次
- Explaining the Ambiguity of Object Detection and 6D Pose From Visual DataFabian Manhardt, Diego Martín Arroyo, Christian Rupprecht, Benjamin Busam 等ICCV 2019 · 被引用 139 次
- Sketch3T: Test-Time Training for Zero-Shot SBIRAneeshan Sain, Ayan Kumar Bhunia, Vaishnav Potlapalli, Pinaki Nath Chowdhury 等CVPR 2022 · 被引用 55 次
相关 Paper
- Data-Free Sketch-Based Image RetrievalAbhra Chaudhuri, Ayan Kumar Bhunia, Yi-Zhe Song, Anjan DuttaCVPR 2023
- Sketch2Mesh: Reconstructing and Editing 3D Shapes from SketchesBenoît Guillard, Edoardo Remelli, Pierre Yvernay, Pascal FuaICCV 2021 · 被引用 102 次
- Exploiting Unlabelled Photos for Stronger Fine-Grained SBIRAneeshan Sain, Ayan Kumar Bhunia, Subhadeep Koley, Pinaki Nath Chowdhury 等CVPR 2023
- Sketch2Model: View-Aware 3D Modeling From Single Free-Hand SketchesSong-Hai Zhang, Yuan-Chen Guo, Qing-Wen GuCVPR 2021
- CLIP for All Things Zero-Shot Sketch-Based Image Retrieval, Fine-Grained or NotAneeshan Sain, Ayan Kumar Bhunia, Pinaki Nath Chowdhury, Subhadeep Koley 等CVPR 2023
