UrbanBIS: a Large-scale Benchmark for Fine-grained Urban Building Instance Segmentation
Guoqing Yang, Fuyou Xue, Qi Zhang, Ke Xie, Chi-Wing Fu, Hui Huang
Abstract
We present the UrbanBIS benchmark for large-scale 3D urban understanding, supporting practical urban-level semantic and building-level instance segmentation. UrbanBIS comprises six real urban scenes, with 2.5 billion points, covering a vast area of 10.78 km2 and 3,370 buildings, captured by 113,346 views of aerial photogrammetry. Particularly, UrbanBIS provides not only semantic-level annotations on a rich set of urban objects, including buildings, vehicles, vegetation, roads, and bridges, but also instance-level annotations on the buildings. Further, UrbanBIS is the first 3D dataset that introduces fine-grained building sub-categories, considering a wide variety of shapes for different building types. Besides, we propose B-Seg, a building instance segmentation method to establish UrbanBIS. B-Seg adopts an end-to-end framework with a simple yet effective strategy for handling large-scale point clouds. Compared with mainstream methods, B-Seg achieves better accuracy with faster inference speed on UrbanBIS. In addition to the carefully-annotated point clouds, UrbanBIS provides high-resolution aerial-acquisition photos and high-quality large-scale 3D reconstruction models, which shall facilitate a wide range of studies such as multi-view stereo, urban LOD generation, aerial path planning, autonomous navigation, road network extraction, and so on, thus serving as an important platform for many intelligent city applications. UrbanBIS and related code can be downloaded at https://vcc.tech/UrbanBIS.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- 3D Question Answering for City Scene UnderstandingPenglei Sun, Yaoxian Song, Xiang Liu, Xiaofei Yang et al.ACM MM 2024 · 6 citations
- Aerial Path Online Planning for Urban Scene UpdationMingfeng Tang, Ningna Wang, Ziyuan Xie, Jianwei Hu et al.SIGGRAPH 2025 · 4 citations
- Weighted Poisson-disk Resampling on Large-Scale Point CloudsXianhe Jiao, Chenlei Lv, Junli Zhao, Ran Yi et al.AAAI 2025 · 4 citations
- Aerial Lifting: Neural Urban Semantic and Building Instance Lifting from Aerial ImageryYuqi Zhang, Guanying Chen, Jiaxing Chen, Shuguang CuiCVPR 2024 · 4 citations
- City-VLM: Towards Multidomain Perception Scene Understanding via Multimodal Incomplete LearningPenglei Sun, Yaoxian Song, Xiangru Zhu, Xiang Liu et al.ACM MM 2025 · 2 citations
Builds on16
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- Revisiting Point Cloud Classification: A New Benchmark Dataset and Classification Model on Real-World DataMikaela Angelina Uy, Quang-Hieu Pham, Binh-Son Hua, Duc Thanh Nguyen et al.ICCV 2019 · 1,003 citations
- Hypersim: A Photorealistic Synthetic Dataset for Holistic Indoor Scene UnderstandingMike Roberts, Jason Ramapuram, Anurag Ranjan, Atulit Kumar et al.ICCV 2021 · 633 citations
- SoftGroup for 3D Instance Segmentation on Point CloudsThang Vu, Kookhoi Kim, Tung Minh Luu, Thanh Xuan Nguyen et al.CVPR 2022 · 251 citations
- Hierarchical Aggregation for 3D Instance SegmentationShaoyu Chen, Jiemin Fang, Qian Zhang, Wenyu Liu et al.ICCV 2021 · 211 citations
Related papers
- Towards Semantic Segmentation of Urban-Scale 3D Point Clouds: A Dataset, Benchmarks and ChallengesQingyong Hu, Bo Yang, Sheikh Khalid, Wen Xiao et al.CVPR 2021
- Campus3D: A Photogrammetry Point Cloud Benchmark for Hierarchical Understanding of Outdoor SceneXinke Li, Chongshou Li, Zekun Tong, Andrew Lim et al.ACM MM 2020 · 63 citations
- Building3D: An Urban-Scale Dataset and Benchmarks for Learning Roof Structures from Point CloudsRuisheng Wang, Shangfeng Huang, Hongxin YangICCV 2023 · 55 citations
- Learnable Earth Parser: Discovering 3D Prototypes in Aerial ScansRomain Loiseau, Elliot Vincent, Mathieu Aubry, Loïc LandrieuCVPR 2024 · 3 citations
- CULTURE3D: A Large-Scale and Diverse Dataset of Cultural Landmarks and Terrains for Gaussian-Based Scene RenderingXinyi Zheng, Steve Zhang, Weizhe Lin, Aaron Zhang et al.ICCV 2025 · 2 citations
