ReST: A Reconfigurable Spatial-Temporal Graph Model for Multi-Camera Multi-Object Tracking
Cheng-Che Cheng, Min-Xuan Qiu, Chen-Kuo Chiang, Shang-Hong Lai
Abstract
Multi-Camera Multi-Object Tracking (MC-MOT) utilizes information from multiple views to better handle problems with occlusion and crowded scenes. Recently, the use of graph-based approaches to solve tracking problems has become very popular. However, many current graphbased methods do not effectively utilize information regarding spatial and temporal consistency. Instead, they rely on single-camera trackers as input, which are prone to fragmentation and ID switch errors. In this paper, we propose a novel reconfigurable graph model that first associates all detected objects across cameras spatially before reconfiguring it into a temporal graph for Temporal Association. This two-stage association approach enables us to extract robust spatial and temporal-aware features and address the problem with fragmented tracklets. Furthermore, our model is designed for online tracking, making it suitable for real-world applications. Experimental results show that the proposed graph model is able to extract more discriminating features for object tracking, and our model achieves state-of-the-art performance on several public datasets. Code is available at https://github. com/chengche6230/ReST .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0ada4822-d981-4bed-878c-502f8f21ccffCited by top-tier papers6
- Cross-View Referring Multi-Object TrackingSijia Chen, En Yu, Wenbing TaoAAAI 2025 · 15 citations
- MVTrajecter: Multi-View Pedestrian Tracking With Trajectory Motion Cost and Trajectory Appearance CostTaiga Yamane, Ryo Masumura, Satoshi Suzuki, Shota OrihashiICCV 2025 · 2 citations
- Learning from Synchronization: Self-Supervised Uncalibrated Multi-View Person Association in Challenging ScenesKeqi Chen, Vinkle Srivastav, Didier Mutter, Nicolas PadoyCVPR 2025
- Multi-view Crowd Tracking Transformer with View-Ground Interactions Under Large Real-World ScenesQi Zhang, Jixuan Chen, Zhang Kaiyi, Xinquan Yu et al.CVPR 2026
- ADA-Track: End-to-End Multi-Camera 3D Multi-Object Tracking with Alternating Detection and AssociationShuxiao Ding, Lukas Schneider, Marius Cordts, Juergen GallCVPR 2024
Builds on9
- Tracking Without Bells and WhistlesPhilipp Bergmann, Tim Meinhardt, Laura Leal-TaixéICCV 2019 · 1,030 citations
- Omni-Scale Feature Learning for Person Re-IdentificationKaiyang Zhou, Yongxin Yang, Andrea Cavallaro, Tao XiangICCV 2019 · 997 citations
- FAMNet: Joint Learning of Feature, Affinity and Multi-Dimensional Assignment for Online Multiple Object TrackingPeng Chu, Haibin LingICCV 2019 · 229 citations
- Multiview Detection with Shadow Transformer (and View-Coherent Data Augmentation)Yunzhong Hou, Liang ZhengACM MM 2021 · 65 citations
- LMGP: Lifted Multicut Meets Geometry Projections for Multi-Camera Multi-Object TrackingDuy M. H. Nguyen, Roberto Henschel, Bodo Rosenhahn, Daniel Sonntag et al.CVPR 2022 · 49 citations
Related papers
- DyGLIP: A Dynamic Graph Model With Link Prediction for Accurate Multi-Camera Multiple Object TrackingKha Gia Quach, Pha A. Nguyen, Huu Le, Thanh-Dat Truong et al.CVPR 2021
- STAR: Spatial-Temporal Tracklet Matching for Multi-Object TrackingXuewei Bai, Yongcai Wang, Deying Li, Haodi Ping et al.NeurIPS 2025
- GMT: Effective Global Framework for Multi-Camera Multi-Target TrackingYihao Zhen, Mingyue Xu, Qiang Wang, Baojie Fan et al.CVPR 2026
- MotionTrack: Learning Robust Short-Term and Long-Term Motions for Multi-Object TrackingZheng Qin, Sanping Zhou, Le Wang, Jinghai Duan et al.CVPR 2023
- Learning Multi-Granular Hypergraphs for Video-Based Person Re-IdentificationYichao Yan, Jie Qin, Jiaxin Chen, Li Liu et al.CVPR 2020
