CroCoDL: Cross-device Collaborative Dataset for Localization
Hermann Blum, Alessandro Mercurio, Joshua O'Reilly, Tim Engelbracht, Mihai Dusmanu, Marc Pollefeys, Zuria Bauer
Abstract
Accurate localization plays a pivotal role in the autonomy of systems operating in unfamiliar environments, particularly when interaction with humans is expected. High-accuracy visual localization systems encompass various components, such as image retrievers, feature extractors, matchers, reconstruction and pose estimation methods. This complexity translates to the necessity of robust evaluation settings and pipelines. However, existing datasets and benchmarks primarily focus on single-agent scenarios, overlooking the critical issue of cross-device localization. Different agents with different sensors will show their own specific strengths and weaknesses, and the data they have available varies substantially. This work addresses this gap by enhancing an existing augmented reality visual localization benchmark with data from legged robots, and evaluating human-robot, cross-device mapping and localization. Our contributions extend beyond device diversity and include high environment variability, spanning ten distinct locations ranging from disaster sites to art exhibitions. Each scene in our dataset features recordings from robot agents, hand-held and head-mounted devices, and highaccuracy ground truth LiDAR scanners, resulting in a comprehensive multi-agent dataset and benchmark. This work represents a significant advancement in the field of visual localization benchmarking, with key insights into the performance of cross-device localization methods across diverse settings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on8
- Learning With Average Precision: Training Image Retrieval With a Listwise LossJérôme Revaud, Jon Almazán, Rafael S. Rezende, César Roberto de SouzaICCV 2019 · 424 citations
- Rethinking Visual Geo-localization for Large-Scale ApplicationsGabriele Moreno Berton, Carlo Masone, Barbara CaputoCVPR 2022 · 235 citations
- GLACE: Global Local Accelerated Coordinate EncodingFangjinhua Wang, Xudong Jiang, Silvano Galliani, Christoph Vogel et al.CVPR 2024 · 24 citations
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora et al.CVPR 2020
- Accelerated Coordinate Encoding: Learning to Relocalize in Minutes Using RGB and PosesEric Brachmann, Tommaso Cavallari, Victor Adrian PrisacariuCVPR 2023
Related papers
- CrowdDriven: A New Challenging Dataset for Outdoor Visual LocalizationAra Jafarzadeh, Manuel López-Antequera, Pau Gargallo, Yubin Kuang et al.ICCV 2021 · 18 citations
- 360Loc: A Dataset and Benchmark for Omnidirectional Visual Localization with Cross-Device QueriesHuajian Huang, Changkun Liu, Yipeng Zhu, Hui Cheng et al.CVPR 2024
- Large-Scale Localization Datasets in Crowded Indoor SpacesDonghwan Lee, Soo-Hyun Ryu, Suyong Yeon, Yonghan Lee et al.CVPR 2021
- From Motion to Localization: Cross-view Optimization with Stationary Event and RGB Cameras for Enhanced Pose EstimationYukun Zhao, Xinyuan Song, Huajian Huang, Tristan BraudUbiComp 2026
- Visual Localization using Imperfect 3D Models from the InternetVojtech Panek, Zuzana Kukelova, Torsten SattlerCVPR 2023
