OmniCity: Omnipotent City Understanding with Multi-Level and Multi-View Images
Weijia Li, Yawen Lai, Linning Xu, Yuanbo Xiangli, Jinhua Yu, Conghui He, Gui-Song Xia, Dahua Lin
Abstract
pixel-wise annotation efforts, we propose an efficient streetview image annotation pipeline that leverages the existing label maps of satellite view and the transformation relations between different views (satellite, panorama, and mono-view). With the new OmniCity dataset, we provide benchmarks for a variety of tasks including building footprint extraction, height estimation, and building plane/instance/fine-grained segmentation. Compared with existing multi-level and multi-view benchmarks, OmniCity contains a larger number of images with richer annotation types and more views, provides more benchmark results This CVPR paper is the Open Access version, provided by the Computer Vision Foundation. Except for this watermark, it is identical to the accepted version; the final published version of the proceedings is available on IEEE Xplore.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d5fa914d-126d-4a9a-855d-c36fe4f39e10Cited by top-tier papers14
- Earth-Agent: Unlocking the Full Landscape of Earth Observation with AgentsPeilin Feng, Zhutao Lv, Junyan Ye, Xiaolei Wang et al.ICLR 2026 · 49 citations
- 3D Building Reconstruction from Monocular Remote Sensing ImagesWeijia Li, Lingxuan Meng, Jinwang Wang, Conghui He et al.ICCV 2021 · 46 citations
- CityDreamer: Compositional Generative Model of Unbounded 3D CitiesHaozhe Xie, Zhaoxi Chen, Fangzhou Hong, Ziwei LiuCVPR 2024 · 35 citations
- UrBench: A Comprehensive Benchmark for Evaluating Large Multimodal Models in Multi-View Urban ScenariosBaichuan Zhou, Haote Yang, Dairong Chen, Junyan Ye et al.AAAI 2025 · 34 citations
- SG-BEV: Satellite-Guided BEV Fusion for Cross-View Semantic SegmentationJunyan Ye, Qiyan Luo, Jinhua Yu, Huaping Zhong et al.CVPR 2024 · 19 citations
Builds on17
- CARAFE: Content-Aware ReAssembly of FEaturesJiaqi Wang, Kai Chen, Rui Xu, Ziwei Liu et al.ICCV 2019 · 842 citations
- Bridging the Domain Gap for Ground-to-Aerial Image MatchingKrishna Regmi, Mubarak ShahICCV 2019 · 191 citations
- Rope3D: The Roadside Perception Dataset for Autonomous Driving and Monocular 3D Object Detection TaskXiaoqing Ye, Mao Shu, Hanyu Li, Yifeng Shi et al.CVPR 2022 · 130 citations
- SpaceNet MVOI: A Multi-View Overhead Imagery DatasetNicholas Weir, David Lindenbaum, Alexei Bastidas, Adam Van Etten et al.ICCV 2019 · 79 citations
- SkyScapes - Fine-Grained Semantic Understanding of Aerial ScenesSeyed Majid Azimi, Corentin Henry, Lars Sommer, Arne Schumann et al.ICCV 2019 · 73 citations
Related papers
- UrbanBIS: a Large-scale Benchmark for Fine-grained Urban Building Instance SegmentationGuoqing Yang, Fuyou Xue, Qi Zhang, Ke Xie et al.SIGGRAPH 2023 · 39 citations
- Holistic Multi-View Building Analysis in the Wild with Projection PoolingZbigniew Wojna, Krzysztof Maziarz, Lukasz Jocz, Robert Paluba et al.AAAI 2021 · 7 citations
- Revisiting the Necessity of Full Accuracy: Weakly Supervised Object-Level Offset Correction for Misaligned Building LabelsJunda Xu, Yanmeng Liu, Xiangqiang Zeng, Jinrong Wu et al.CVPR 2026
- Instruction-guided Multi-Granularity Segmentation and Captioning with Large Multimodal ModelXu Yuan, Li Zhou, Zenghui Sun, Zikun Zhou et al.AAAI 2025 · 1 citation
- Multi-View Pedestrian Occupancy Prediction with a Novel Synthetic DatasetSithu Aung, Min-Cheol Sagong, Junghyun ChoAAAI 2025 · 5 citations
