Homography Guided Temporal Fusion for Road Line and Marking Segmentation
Shan Wang, Chuong Nguyen, Jiawei Liu, Kaihao Zhang, Wenhan Luo, Yanhao Zhang, Sundaram Muthu, Fahira Afzal Maken, Hongdong Li
Abstract
Reliable segmentation of road lines and markings is critical to autonomous driving. Our work is motivated by the observations that road lines and markings are (1) frequently occluded in the presence of moving vehicles, shadow, and glare and (2) highly structured with low intra-class shape variance and overall high appearance consistency. To solve these issues, we propose a Homography Guided Fusion (HomoFusion) module to exploit temporally-adjacent video frames for complementary cues facilitating the correct classification of the partially occluded road lines or markings. To reduce computational complexity, a novel surface normal estimator is proposed to establish spatial correspondences between the sampled frames, allowing the Homo-Fusion module to perform a pixel-to-pixel attention mechanism in updating the representation of the occluded road lines or markings. Experiments on ApolloScape, a largescale lane mark segmentation dataset, and ApolloScape Night with artificial simulated night-time road conditions, demonstrate that our method outperforms other existing SOTA lane mark segmentation models with less than 9% of their parameters and computational complexity. We show that exploiting available camera intrinsic data and ground plane assumption for cross-frame correspondence can lead to a light-weight network with significantly improved performances in speed and accuracy. We also prove the versatility of our HomoFusion approach by applying it to the problem of water puddle segmentation and achieving SOTA performance 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cf9c4e83-1ff1-48f3-bf76-3a0225f6bfc6Cited by top-tier papers1
Ask how each one uses itBuilds on14
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Large Batch Optimization for Deep Learning: Training BERT in 76 minutesYang You, Jing Li, Sashank J. Reddi, Jonathan Hseu et al.ICLR 2020 · 1,170 citations
- PETRv2: A Unified Framework for 3D Perception from Multi-Camera ImagesYingfei Liu, Junjie Yan, Fan Jia, Shuailin Li et al.ICCV 2023 · 513 citations
- CondLaneNet: a Top-to-down Lane Detection Framework Based on Conditional ConvolutionLizhe Liu, Xiaohao Chen, Siyu Zhu, Ping TanICCV 2021 · 312 citations
- VIL-100: A New Dataset and A Baseline Model for Video Instance Lane DetectionYujun Zhang, Lei Zhu, Wei Feng, Huazhu Fu et al.ICCV 2021 · 67 citations
Related papers
- A Hybrid Global-Local Perception Network for Lane DetectionQing Chang, Yifei TongAAAI 2024 · 14 citations
- Vanishing-Point-Guided Video Semantic Segmentation of Driving ScenesDiandian Guo, Deng-Ping Fan, Tongyu Lu, Christos Sakaridis et al.CVPR 2024
- Focus on Local: Detecting Lane Marker From Bottom Up via Key PointZhan Qu, Huan Jin, Yang Zhou, Zhen Yang et al.CVPR 2021
- LaneSegNet: Map Learning with Lane Segment Perception for Autonomous DrivingTianyu Li, Peijin Jia, Bangjun Wang, Li Chen et al.ICLR 2024 · 69 citations
- Keep Your Eyes on the Lane: Real-Time Attention-Guided Lane DetectionLucas Tabelini Torres, Rodrigo Ferreira Berriel, Thiago M. Paixão, Claudine Badue et al.CVPR 2021
