CTRL-C: Camera calibration TRansformer with Line-Classification
Jinwoo Lee, Hyunsung Go, Hyunjoon Lee, Sunghyun Cho, Min-Hyuk Sung, Junho Kim
Abstract
Single image camera calibration is the task of estimating the camera parameters from a single input image, such as the vanishing points, focal length, and horizon line. In this work, we propose Camera calibration TRansformer with Line-Classification (CTRL-C), an end-to-end neural network-based approach to single image camera calibration, which directly estimates the camera parameters from an image and a set of line segments. Our network adopts the transformer architecture to capture the global structure of an image with multi-modal inputs in an end-to-end manner. We also propose an auxiliary task of line classification to train the network to extract the global geometric information from lines effectively. Our experiments demonstrate that CTRL-C outperforms the previous stateof-the-art methods on the Google Street View and SUN360 benchmark datasets. Code is available at https:// github . com/ jwlee-vcl/ CTRL-C.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 10737bf9-e544-45e0-9bf2-e215ce6d7e1fCited by top-tier papers14
- Future Transformer for Long-term Action AnticipationDayoung Gong, Joonseok Lee, Manjin Kim, Seong Jong Ha et al.CVPR 2022 · 56 citations
- Tame a Wild Camera: In-the-Wild Monocular Camera CalibrationShengjie Zhu, Abhinav Kumar, Masa Hu, Xiaoming LiuNeurIPS 2023 · 47 citations
- Deep geometry-aware camera self-calibration from videoAnnika Hagemann, Moritz Knorr, Christoph StillerICCV 2023 · 33 citations
- Collaborative Transformers for Grounded Situation RecognitionJunhyeong Cho, Youngseok Yoon, Suha KwakCVPR 2022 · 23 citations
- Thinking with Camera: A Unified Multimodal Model for Camera-Centric Understanding and GenerationKang Liao, Size Wu, Zhonghua Wu, Linyi Jin et al.ICLR 2026 · 19 citations
Builds on5
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- UprightNet: Geometry-Aware Camera Orientation Estimation From Single ImagesWenqi Xian, Zhengqi Li, Noah Snavely, Matthew Fisher et al.ICCV 2019 · 52 citations
- 12-in-1: Multi-Task Vision and Language Representation LearningJiasen Lu, Vedanuj Goswami, Marcus Rohrbach, Devi Parikh et al.CVPR 2020
- Line Segment Detection Using Transformers Without EdgesYifan Xu, Weijian Xu, David Cheung, Zhuowen TuCVPR 2021
Related papers
- End-to-End Real-Time Vanishing Point Detection with TransformerXin Tong, Shi Peng, Yufei Guo, Xuhui HuangAAAI 2024 · 4 citations
- End-to-End Camera Calibration for Broadcast VideosLong Sha, Jennifer A. Hobbs, Panna Felsen, Xinyu Wei et al.CVPR 2020
- Learning Multi-Scene Absolute Pose Regression with TransformersYoli Shavit, Ron Ferens, Yosi KellerICCV 2021 · 163 citations
- Transformer Based Line Segment Classifier with Image Context for Real-Time Vanishing Point Detection in Manhattan WorldXin Tong, Xianghua Ying, Yongjie Shi, Ruibin Wang et al.CVPR 2022 · 17 citations
- AnyCalib: On-Manifold Learning for Model-Agnostic Single-View Camera CalibrationJavier Tirado-Garín, Javier CiveraICCV 2025 · 9 citations
