SPARE3D: A Dataset for SPAtial REasoning on Three-View Line Drawings
Wenyu Han, Siyuan Xiang, Chenhui Liu, Ruoyu Wang, Chen Feng
Abstract
Spatial reasoning is an important component of human intelligence. We can imagine the shapes of 3D objects and reason about their spatial relations by merely looking at their three-view line drawings in 2D, with different levels of competence. Can deep networks be trained to perform spatial reasoning tasks? How can we measure their "spatial intelligence"? To answer these questions, we present the SPARE3D dataset. Based on cognitive science and psychometrics, SPARE3D contains three types of 2D-3D reasoning tasks on view consistency, camera pose, and shape generation, with increasing difficulty. We then design a method to automatically generate a large number of challenging questions with ground truth answers for each task. They are used to provide supervision for training our baseline models using state-of-the-art architectures like ResNet. Our experiments show that although convolutional networks have achieved superhuman performance in many visual learning tasks, their spatial reasoning performance in SPARE3D is almost equal to random guesses. We hope SPARE3D can stimulate new problem formulations and network designs for spatial reasoning to empower intelligent robots to operate effectively in the 3D world via 2D sensors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e38d2f77-85a3-4d0b-bd60-f49bc08a8cf2Cited by top-tier papers9
- Brick-by-Brick: Combinatorial Construction with Deep Reinforcement LearningHyunsoo Chung, Jungtaek Kim, Boris Knyazev, Jinhwi Lee et al.NeurIPS 2021 · 29 citations
- From 2D CAD Drawings to 3D Parametric Models: A Vision-Language ApproachXilin Wang, Jia Zheng, Yuanchao Hu, Hao Zhu et al.AAAI 2025 · 14 citations
- PlankAssembly: Robust 3D Reconstruction from Three Orthographic Views with Learnt Shape ProgramsWentao Hu, Jia Zheng, Zixin Zhang, Xiaojun Yuan et al.ICCV 2023 · 12 citations
- Freehand Sketch Generation from Mechanical ComponentsZhichao Liao, Fengyuan Piao, Di Huang, Xinghui Li et al.ACM MM 2024 · 12 citations
- CReFT-CAD: Boosting Orthographic Projection Reasoning for CAD via Reinforcement Fine-TuningKe Niu, Zhuofan Chen, Haiyang Yu, Yuwen Chen et al.NeurIPS 2025 · 8 citations
Builds on1
Related papers
- Self-supervised Spatial Reasoning on Multi-View Line DrawingsSiyuan Xiang, Anbang Yang, Yanfei Xue, Yaoqing Yang et al.CVPR 2022 · 4 citations
- 3DSRBENCH: A Comprehensive 3D Spatial Reasoning BenchmarkWufei Ma, Haoyu Chen, Guofeng Zhang, Yu-Cheng Chou et al.ICCV 2025 · 15 citations
- SITE: Towards Spatial Intelligence Thorough EvaluationWenqi Wang, Reuben Tan, Pengyue Zhu, Jianwei Yang et al.ICCV 2025 · 2 citations
- SpaRE: Enhancing Spatial Reasoning in Vision-Language Models with Synthetic DataMichael Ogezi, Freda ShiACL 2025 · 19 citations
- Learning Multi-View Spatial Reasoning from Cross-View RelationsSuchae Jeong, Jaehwi Song, Haeone Lee, Hanna Kim et al.CVPR 2026
