MultiCam: On-the-fly Multi-Camera Pose Estimation Using Spatiotemporal Overlaps of Known Objects
Shiyu Li, Hannah Schieber, Kristoffer Waldow, Benjamin Busam, Julian Kreimeier, Daniel Roth
Abstract
Multi-camera dynamic Augmented Reality (AR) applications require a camera pose estimation to leverage individual information from each camera in one common system. This can be achieved by combining contextual information, such as markers or objects, across multiple views. While commonly cameras are calibrated in an initial step or updated through the constant use of markers, another option is to leverage information already present in the scene, like known objects. Another downside of marker-based tracking is that markers have to be tracked inside the field-of-view (FoV) of the cameras. To overcome these limitations, we propose a constant dynamic camera pose estimation leveraging spatiotemporal FoV overlaps of known objects on the fly. To achieve that, we enhance the state-of-the-art object pose estimator to update our spatiotemporal scene graph, enabling a relation even among non-overlapping FoV cameras. To evaluate our approach, we introduce a multi-camera, multi-object pose estimation dataset with temporal FoV overlap, including static and dynamic cameras. Furthermore, in FoV overlapping scenarios, we outperform the state-of-the-art on the widely used YCB-V and T-LESS dataset in camera pose accuracy. Our performance on both previous and our proposed datasets validates the effectiveness of our marker-less approach for AR applications. The code and dataset are available on https://github.com/roth-hex-lab/IEEE-VR-2026-MultiCam.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3c3bf150-75ec-41df-b3e9-5d04529894cbBuilds on20
- BARF: Bundle-Adjusting Neural Radiance FieldsChen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, Simon LuceyICCV 2021 · 867 citations
- RIO: 3D Object Instance Re-Localization in Changing Indoor EnvironmentsJohanna Wald, Armen Avetisyan, Nassir Navab, Federico Tombari et al.ICCV 2019 · 233 citations
- Learning Multi-Scene Absolute Pose Regression with TransformersYoli Shavit, Ron Ferens, Yosi KellerICCV 2021 · 163 citations
- GigaPose: Fast and Robust Novel Object Pose Estimation via One CorrespondenceVan Nguyen Nguyen, Thibault Groueix, Mathieu Salzmann, Vincent LepetitCVPR 2024 · 67 citations
- RTMO: Towards High-Performance One-Stage Real-Time Multi-Person Pose EstimationPeng Lu, Tao Jiang, Yining Li, Xiangtai Li et al.CVPR 2024 · 66 citations
Related papers
- GBOT: Graph-Based 3D Object Tracking for Augmented Reality-Assisted Assembly GuidanceShiyu Li, Hannah Schieber, Niklas Corell, Bernhard Egger et al.IEEE VR 2024 · 25 citations
- MV-TAP: Tracking Any Point in Multi-View VideosJahyeok Koo, Inès Hyeonsu Kim, Mungyeom Kim, Junghyun Park et al.CVPR 2026 · 4 citations
- EgoPoseVR: Spatiotemporal Multi-Modal Reasoning for Egocentric Full-Body Pose in Virtual RealityHaojie Cheng, Shaun Jing Heng Ong, Shaoyu Cai, Aiden Tat Yang Koh et al.IEEE VR 2026 · 1 citation
- BCOT: A Markerless High-Precision 3D Object Tracking BenchmarkJiachen Li, Bin Wang, Shiqiang Zhu, Xin Cao et al.CVPR 2022 · 15 citations
- Cross-View Tracking for Multi-Human 3D Pose Estimation at Over 100 FPSLong Chen, Haizhou Ai, Rui Chen, Zijie Zhuang et al.CVPR 2020
