End-to-End Real-Time Vanishing Point Detection with Transformer
Xin Tong, Shi Peng, Yufei Guo, Xuhui Huang
Abstract
In this paper, we propose a novel transformer-based endto-end real-time vanishing point detection method, which is named Vanishing Point TRansformer (VPTR). The proposed method can directly regress the locations of vanishing points from given images. To achieve this goal, we pose vanishing point detection as a point object detection task on the Gaussian hemisphere with region division. Considering low-level features always provide more geometric information which can contribute to accurate vanishing point prediction, we propose a clear architecture where vanishing point queries in the decoder can directly gather multi-level features from CNN backbone with deformable attention in VPTR. Our method does not rely on line detection or Manhattan world assumption, which makes it more flexible to use. VPTR runs at an inferring speed of 140 FPS on one NVIDIA 3090 card. Experimental results on synthetic and real-world datasets demonstrate that our method can be used in both natural and structural scenes, and is superior to other state-of-the-art methods on the balance of accuracy and efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Improving Transformer Based Line Segment Detection with Matched Predicting and Re-rankingXin Tong, Shi Peng, Baojie Tian, Yufei Guo et al.AAAI 2025 · 3 citations
- RANK++LETR: Learn to Rank and Optimize Candidates for Line Segment DetectionXin Tong, Baojie Tian, Yufei Guo, Zhe MaNeurIPS 2025
Builds on12
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- Learning to Reconstruct 3D Manhattan Wireframes From a Single ImageYichao Zhou, Haozhi Qi, Yuexiang Zhai, Qi Sun et al.ICCV 2019 · 74 citations
- Quasi-Globally Optimal and Efficient Vanishing Point Estimation in Manhattan WorldHaoang Li, Ji Zhao, Jean-Charles Bazin, Wen Chen et al.ICCV 2019 · 34 citations
- Deep vanishing point detection: Geometric priors make dataset variations vanishYancong Lin, Ruben Wiersma, Silvia L. Pintea, Klaus Hildebrandt et al.CVPR 2022 · 24 citations
Related papers
- VPDETR: End-to-End Vanishing Point DEtection TRansformersTaiyan Chen, Xianghua Ying, Jinfa Yang, Ruibin Wang et al.AAAI 2024 · 5 citations
- Transformer Based Line Segment Classifier with Image Context for Real-Time Vanishing Point Detection in Manhattan WorldXin Tong, Xianghua Ying, Yongjie Shi, Ruibin Wang et al.CVPR 2022 · 17 citations
- VaPiD: A Rapid Vanishing Point Detector via Learned OptimizersShichen Liu, Yichao Zhou, Yajie ZhaoICCV 2021 · 19 citations
- DETRs Beat YOLOs on Real-time Object DetectionYian Zhao, Wenyu Lv, Shangliang Xu, Jinman Wei et al.CVPR 2024 · 3,046 citations
- End-to-End Video Object Detection with Spatial-Temporal TransformersLu He, Qianyu Zhou, Xiangtai Li, Li Niu et al.ACM MM 2021 · 106 citations
