BCOT: A Markerless High-Precision 3D Object Tracking Benchmark
Jiachen Li, Bin Wang, Shiqiang Zhu, Xin Cao, Fan Zhong, Wenxuan Chen, Te Li, Jason Gu, Xueying Qin
Abstract
Template-based 3D object tracking still lacks a high-precision benchmark of real scenes due to the difficulty of annotating the accurate 3D poses of real moving video objects without using markers. In this paper, we present a multi-view approach to estimate the accurate 3D poses of real moving objects, and then use binocular data to construct a new benchmark for monocular textureless 3D object tracking. The proposed method requires no markers, and the cameras only need to be synchronous, relatively fixed as cross-view and calibrated. Based on our object-centered model, we jointly optimize the object pose by minimizing shape reprojection constraints in all views, which greatly improves the accuracy compared with the single-view approach, and is even more accurate than the depth-based method. Our new benchmark dataset contains 20 textureless objects, 22 scenes, 404 video sequences and 126K images captured in real scenes. The annotation error is guaranteed to be less than 2mm, according to both theoretical analysis and validation experiments. We reevaluate the state-of-the-art 3D object tracking methods with our dataset, reporting their performance ranking in real scenes. Our BCOT benchmark and code can be found at https://ar3dv.github.io/BCOT-Benchmark/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 15e67487-589b-418d-aa57-14201efd9f5aCited by top-tier papers5
- Deep Active Contours for Real-time 6-DoF Object TrackingLong Wang, Shen Yan, Jianan Zhen, Yu Liu et al.ICCV 2023 · 26 citations
- GBOT: Graph-Based 3D Object Tracking for Augmented Reality-Assisted Assembly GuidanceShiyu Li, Hannah Schieber, Niklas Corell, Bernhard Egger et al.IEEE VR 2024 · 25 citations
- MagDesk: Interactive Tabletop Workspace Based on Passive Magnetic TrackingKunpeng Huang, Yasha Iravantchi, Dongyao Chen, Alanson P. SampleUbiComp 2025 · 6 citations
- MultiCam: On-the-fly Multi-Camera Pose Estimation Using Spatiotemporal Overlaps of Known ObjectsShiyu Li, Hannah Schieber, Kristoffer Waldow, Benjamin Busam et al.IEEE VR 2026 · 2 citations
- Globally Optimal Pose from Orthographic SilhouettesAgniva Sengupta, Dilara Kus, Jianning Li, Stefan ZachowCVPR 2026
Builds on4
- StereOBJ-1M: Large-scale Stereo Image Dataset for 6D Object Pose EstimationXingyu Liu, Shun Iwase, Kris M. KitaniICCV 2021 · 58 citations
- GraspNet-1Billion: A Large-Scale Benchmark for General Object GraspingHaoshu Fang, Chenxi Wang, Minghao Gou, Cewu LuCVPR 2020
- GDR-Net: Geometry-Guided Direct Regression Network for Monocular 6D Object Pose EstimationGu Wang, Fabian Manhardt, Federico Tombari, Xiangyang JiCVPR 2021
- KeyPose: Multi-View 3D Labeling and Keypoint Estimation for Transparent ObjectsXingyu Liu, Rico Jonschkowski, Anelia Angelova, Kurt KonoligeCVPR 2020
Related papers
- CDTB: A Color and Depth Visual Object Tracking Dataset and BenchmarkAlan Lukezic, Ugur Kart, Jani Käpylä, Ahmed Durmush et al.ICCV 2019 · 79 citations
- 360VOT: A New Benchmark Dataset for Omnidirectional Visual Object TrackingHuajian Huang, Yinzhe Xu, Yingshu Chen, Sai-Kit YeungICCV 2023 · 12 citations
- Cross-View Tracking for Multi-Human 3D Pose Estimation at Over 100 FPSLong Chen, Haizhou Ai, Rui Chen, Zijie Zhuang et al.CVPR 2020
- Tracking Without Bells and WhistlesPhilipp Bergmann, Tim Meinhardt, Laura Leal-TaixéICCV 2019 · 1,030 citations
- MITracker: Multi-View Integration for Visual Object TrackingMengjie Xu, Yitao Zhu, Haotian Jiang, Jiaming Li et al.CVPR 2025
