Holistic Multi-View Building Analysis in the Wild with Projection Pooling
Zbigniew Wojna, Krzysztof Maziarz, Lukasz Jocz, Robert Paluba, Robert Kozikowski, Iasonas Kokkinos
摘要
We address six different classification tasks related to fine-grained building attributes: construction type, number of floors, pitch and geometry of the roof, facade material, and occupancy class. Tackling such a remote building analysis problem became possible only recently due to growing large-scale datasets of urban scenes. To this end, we introduce a new benchmarking dataset, consisting of 49426 images (top-view and street-view) of 9674 buildings. These photos are further assembled, together with the geometric metadata. The dataset showcases various real-world challenges, such as occlusions, blur, partially visible objects, and a broad spectrum of buildings. We propose a new projection pooling layer, creating a unified, top-view representation of the top-view and the side views in a high-dimensional space. It allows us to utilize the building and imagery metadata seamlessly. Introducing this layer improves classification accuracy -- compared to highly tuned baseline models -- indicating its suitability for building analysis.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- SG-BEV: Satellite-Guided BEV Fusion for Cross-View Semantic SegmentationJunyan Ye, Qiyan Luo, Jinhua Yu, Huaping Zhong 等CVPR 2024 · 被引用 19 次
- OmniCity: Omnipotent City Understanding with Multi-Level and Multi-View ImagesWeijia Li, Yawen Lai, Linning Xu, Yuanbo Xiangli 等CVPR 2023
相关 Paper
- UrbanBIS: a Large-scale Benchmark for Fine-grained Urban Building Instance SegmentationGuoqing Yang, Fuyou Xue, Qi Zhang, Ke Xie 等SIGGRAPH 2023 · 被引用 39 次
- BuildingNet: Learning to Label 3D BuildingsPratheba Selvaraju, Mohamed Nabail, Marios Loizou, Maria Maslioukova 等ICCV 2021 · 被引用 54 次
- SkyScapes - Fine-Grained Semantic Understanding of Aerial ScenesSeyed Majid Azimi, Corentin Henry, Lars Sommer, Arne Schumann 等ICCV 2019 · 被引用 73 次
- An Instance-Centric Panoptic Occupancy Prediction Benchmark for Autonomous DrivingYi Feng, Junwu E, Zizhan Guo, Yu Ma 等CVPR 2026 · 被引用 1 次
- Google Landmarks Dataset v2 - A Large-Scale Benchmark for Instance-Level Recognition and RetrievalTobias Weyand, André Araújo, Bingyi Cao, Jack SimCVPR 2020
