ACPV-Net: All-Class Polygonal Vectorization for Seamless Vector Map Generation from Aerial Imagery
Weiqin Jiao, Hao Cheng, George Vosselman, Claudio Persello
Abstract
We tackle the problem of generating a complete vector map representation from aerial imagery in a single run: producing polygons for all land-cover classes with shared boundaries and no gaps or overlaps. Existing polygonization methods are typically class-specific; extending them to multiple classes via per-class runs commonly leads to topological inconsistencies, such as duplicated edges, gaps, and overlaps. We formalize this new task as All-Class Polygonal Vectorization (ACPV) and release the first public benchmark, Deventer-512, with standardized metrics jointly evaluating semantic fidelity, geometric accuracy, vertex efficiency, per-class topological fidelity and global topological consistency. To realize ACPV, we propose ACPV-Net, a unified framework introducing a novel Semantically Supervised Conditioning (SSC) mechanism coupling semantic perception with geometric primitive generation, along with a topological reconstruction that guarantees shared-edge consistency by design. While enforcing such strict topological constraints, ACPV-Net surpasses all class-specific baselines in polygon quality across classes on Deventer-512, e.g., compared to TopDiG on vegetation, +9.9 IoU (semantic fidelity), -45% PoLiS (geometric error), -59% N-ratio (vertex redundancy). It also applies to single-class polygonal vectorization without any architectural modification, achieving the best-reported results on WHU-Building. Data, code, and models will be publicly released.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 78e126f0-0a75-4d7c-8bc6-7e50f1a6cb24Builds on11
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- VMamba: Visual State Space ModelYue Liu, Yunjie Tian, Yuzhong Zhao, Hongtian Yu et al.NeurIPS 2024 · 3,199 citations
- ViTPose: Simple Vision Transformer Baselines for Human Pose EstimationYufei Xu, Jing Zhang, Qiming Zhang, Dacheng TaoNeurIPS 2022 · 1,105 citations
- Large Selective Kernel Network for Remote Sensing Object DetectionYuxuan Li, Qibin Hou, Zhaohui Zheng, Ming-Ming Cheng et al.ICCV 2023 · 535 citations
Related papers
- Joint Semantic-geometric Learning for Polygonal Building SegmentationWeijia Li, Wenqian Zhao, Huaping Zhong, Conghui He et al.AAAI 2021 · 50 citations
- Topological Map Extraction From Overhead ImagesZuoyue Li, Jan Dirk Wegner, Aurélien LucchiICCV 2019 · 181 citations
- Approximating shapes in images with low-complexity polygonsMuxingzi Li, Florent Lafarge, Renaud MarletCVPR 2020
- Re: PolyWorld - A Graph Neural Network for Polygonal Scene ParsingStefano Zorzi, Friedrich FraundorferICCV 2023 · 12 citations
- Polygonal Building Extraction by Frame Field LearningNicolas Girard, Dmitriy Smirnov, Justin Solomon, Yuliya TarabalkaCVPR 2021
