Towards Light-Weight and Real-Time Line Segment Detection
Geonmo Gu, ByungSoo Ko, SeoungHyun Go, Sung-Hyun Lee, Jingeun Lee, Minchul Shin
摘要
Previous deep learning-based line segment detection (LSD) suffers from the immense model size and high computational cost for line prediction. This constrains them from real-time inference on computationally restricted environments. In this paper, we propose a real-time and light-weight line segment detector for resource-constrained environments named Mobile LSD (M-LSD). We design an extremely efficient LSD architecture by minimizing the backbone network and removing the typical multi-module process for line prediction found in previous methods. To maintain competitive performance with a light-weight network, we present novel training schemes: Segments of Line segment (SoL) augmentation, matching and geometric loss. SoL augmentation splits a line segment into multiple subparts, which are used to provide auxiliary line data during the training process. Moreover, the matching and geometric loss allow a model to capture additional geometric cues. Compared with TP-LSD-Lite, previously the best real-time LSD method, our model (M-LSDtiny) achieves competitive performance with 2.5% of model size and an increase of 130.5% in inference speed on GPU. Furthermore, our model runs at 56.8 FPS and 48.6 FPS on the latest Android and iPhone mobile devices, respectively. To the best of our knowledge, this is the first real-time deep LSD available on mobile devices. Our code is available 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- Uni-ControlNet: All-in-One Control to Text-to-Image Diffusion ModelsShihao Zhao, Dongdong Chen, Yen-Chun Chen, Jianmin Bao 等NeurIPS 2023 · 被引用 505 次
- When ControlNet Meets Inexplicit Masks: A Case Study of ControlNet on its Contour-following AbilityWenjie Xuan, Yufei Xu, Shanshan Zhao, Chaoyue Wang 等ACM MM 2024 · 被引用 3 次
- VisualCloze: A Universal Image Generation Framework via Visual in-Context LearningZhong-Yu Li, Ruoyi Du, Juncheng Yan, Le Zhuo 等ICCV 2025 · 被引用 2 次
- SceneLoom: Communicating Data with Scene ContextLin Gao, Leixian Shen, Yuheng Zhao, Jiexiang Lan 等IEEE VIS 2025 · 被引用 1 次
它引用的顶会 Paper4
- End-to-End Wireframe ParsingYichao Zhou, Haozhi Qi, Yi MaICCV 2019 · 被引用 190 次
- LGNN: A Context-aware Line Segment DetectorQuan Meng, Jiakai Zhang, Qiang Hu, Xuming He 等ACM MM 2020 · 被引用 29 次
- Line Segment Detection Using Transformers Without EdgesYifan Xu, Weijian Xu, David Cheung, Zhuowen TuCVPR 2021
- Holistically-Attracted Wireframe ParsingNan Xue, Tianfu Wu, Song Bai, Fudong Wang 等CVPR 2020
相关 Paper
- ScaleLSD: Scalable Deep Line Segment Detection StreamlinedZeran Ke, Bin Tan, Xianwei Zheng, Yujun Shen 等CVPR 2025
- DeepLSD: Line Segment Detection and Refinement with Deep Image GradientsRémi Pautrat, Daniel Barath, Viktor Larsson, Martin R. Oswald 等CVPR 2023
- SOLD2: Self-Supervised Occlusion-Aware Line Description and DetectionRémi Pautrat, Juan-Ting Lin, Viktor Larsson, Martin R. Oswald 等CVPR 2021
- ELSD: Efficient Line Segment Detector and DescriptorHaotian Zhang, Yicheng Luo, Fangbo Qin, Yijia He 等ICCV 2021 · 被引用 32 次
- SYENet: A Simple Yet Effective Network for Multiple Low-Level Vision Tasks with Real-time Performance on Mobile DeviceWeiran Gou, Ziyao Yi, Yan Xiang, Shaoqing Li 等ICCV 2023 · 被引用 12 次
