Joint Semantic-geometric Learning for Polygonal Building Segmentation
Weijia Li, Wenqian Zhao, Huaping Zhong, Conghui He, Dahua Lin
摘要
Building extraction from aerial or satellite images has been an important research problem in remote sensing and computer vision domains for decades. Compared with pixel-wise semantic segmentation models that output raster building segmentation map, polygonal building segmentation approaches produce more realistic building polygons that are in the desirable vector format for practical applications. Despite the substantial efforts over recent years, state-of-the-art polygonal building segmentation methods still suffer from several limitations, e.g., (1) relying on a perfect segmentation map to guarantee the vectorization quality; (2) requiring a complex post-processing procedure; (3) generating inaccurate vertices with a fixed quantity, a wrong sequential order, self-intersections, etc. To tackle the above issues, in this paper, we propose a polygonal building segmentation approach and make the following contributions: (1) We design a multitask segmentation network for joint semantic and geometric learning via three tasks, i.e., pixel-wise building segmentation, multi-class corner prediction, and edge orientation prediction. (2) We propose a simple but effective vertex generation module for transforming the segmentation contour into high-quality polygon vertices. (3) We further propose a polygon refinement network that automatically moves the polygon vertices into more accurate locations. Results on two popular building segmentation datasets demonstrate that our approach achieves significant improvements for both building instance segmentation (with 2% F1-score gain) and polygon vertex prediction (with 6% F1-score gain) compared with current state-of-the-art methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- MapTR: Structured Modeling and Learning for Online Vectorized HD Map ConstructionBencheng Liao, Shaoyu Chen, Xinggang Wang, Tianheng Cheng 等ICLR 2023 · 被引用 69 次
- 3D Building Reconstruction from Monocular Remote Sensing ImagesWeijia Li, Lingxuan Meng, Jinwang Wang, Conghui He 等ICCV 2021 · 被引用 46 次
- Data Leakage Detection and De-duplication in Large Scale Geospatial Image DatasetsYeshwanth Kumar Adimoolam, Charalambos Poullis, Melinos AverkiouCVPR 2026
- 3D Building Reconstruction from Monocular Remote Sensing Images with Multi-level SupervisionsWeijia Li, Haote Yang, Zhenghao Hu, Juepeng Zheng 等CVPR 2024
- OmniCity: Omnipotent City Understanding with Multi-Level and Multi-View ImagesWeijia Li, Yawen Lai, Linning Xu, Yuanbo Xiangli 等CVPR 2023
它引用的顶会 Paper5
- Topological Map Extraction From Overhead ImagesZuoyue Li, Jan Dirk Wegner, Aurélien LucchiICCV 2019 · 被引用 181 次
- End to End Trainable Active Contours via Differentiable RenderingShir Gur, Tal Shaharabany, Lior WolfICLR 2020 · 被引用 39 次
- Boundary-Aware 3D Building Reconstruction From a Single Overhead ImageJisan Mahmud, True Price, Akash Bapat, Jan-Michael FrahmCVPR 2020
- Approximating shapes in images with low-complexity polygonsMuxingzi Li, Florent Lafarge, Renaud MarletCVPR 2020
- PolyTransform: Deep Polygon Transformer for Instance SegmentationJustin Liang, Namdar Homayounfar, Wei-Chiu Ma, Yuwen Xiong 等CVPR 2020
相关 Paper
- Polygonal Building Extraction by Frame Field LearningNicolas Girard, Dmitriy Smirnov, Justin Solomon, Yuliya TarabalkaCVPR 2021
- Re: PolyWorld - A Graph Neural Network for Polygonal Scene ParsingStefano Zorzi, Friedrich FraundorferICCV 2023 · 被引用 12 次
- PolyWorld: Polygonal Building Extraction with Graph Neural Networks in Satellite ImagesStefano Zorzi, Shabab Bazrafkan, Stefan Habenschuss, Friedrich FraundorferCVPR 2022 · 被引用 99 次
- ACPV-Net: All-Class Polygonal Vectorization for Seamless Vector Map Generation from Aerial ImageryWeiqin Jiao, Hao Cheng, George Vosselman, Claudio PerselloCVPR 2026 · 被引用 3 次
- CVNet: Contour Vibration Network for Building ExtractionZiqiang Xu, Chunyan Xu, Zhen Cui, Xiangwei Zheng 等CVPR 2022 · 被引用 24 次
