Collaborative Tracking Learning for Frame-Rate-Insensitive Multi-Object Tracking
Yiheng Liu, Junta Wu, Yi Fu
Abstract
Multi-object tracking (MOT) at low frame rates can reduce computational, storage and power overhead to better meet the constraints of edge devices. Many existing MOT methods suffer from significant performance degradation in low-frame-rate videos due to significant location and appearance changes between adjacent frames. To this end, we propose to explore collaborative tracking learning (ColTrack) for frame-rate-insensitive MOT in a query-based end-to-end manner. Multiple historical queries of the same target jointly track it with richer temporal descriptions. Meanwhile, we insert an information refinement module between every two temporal blocking decoders to better fuse temporal clues and refine features. Moreover, a tracking object consistency loss is proposed to guide the interaction between historical queries. Extensive experimental results demonstrate that in high-frame-rate videos, ColTrack obtains higher performance than state-of-the-art methods on large-scale datasets Dancetrack and BDD100K, and outperforms the existing end-to-end methods on MOT17. More importantly, ColTrack has a significant advantage over state-of-the-art methods in low-frame-rate videos, which allows it to obtain faster processing speeds by reducing frame-rate requirements while maintaining higher performance. Code will be released at https://github.com/yolomax/ColTrack
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e6d2c3a4-222e-4da6-aaaf-73df269608a3Cited by top-tier papers5
- SAM2MOT: A Novel Paradigm of Multi-Object Tracking by SegmentationJunjie Jiang, Zelin Wang, Manqi Zhao, Yin Li et al.AAAI 2026 · 19 citations
- Predictive and Near-Optimal Sampling for View Materialization in Video DatabasesYanchao Xu, Dongxiang Zhang, Shuhao Zhang, Sai Wu et al.SIGMOD 2024 · 5 citations
- Dual-Path Temporal Decoder for End-to-End Multi-Object TrackingHyunseop Kim, Juheon Jeong, Hanul Kim, Yeong Jun KohNeurIPS 2025 · 4 citations
- When Trackers Date Fish: A Benchmark and Framework for Underwater Multiple Fish TrackingWeiran Li, Yeqiang Liu, Qiannan Guo, Yijie Wei et al.AAAI 2026 · 1 citation
- GLoMOT: Efficient Online GNN-based Low-Frame-Rate Multi-Object TrackerYaxuan Hu, Jie Hua, Gang Wu, Yuhong Yang et al.AAAI 2026
Builds on11
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- Tracking Without Bells and WhistlesPhilipp Bergmann, Tim Meinhardt, Laura Leal-TaixéICCV 2019 · 1,030 citations
- TrackFormer: Multi-Object Tracking with TransformersTim Meinhardt, Alexander Kirillov, Laura Leal-Taixé, Christoph FeichtenhoferCVPR 2022 · 927 citations
- DanceTrack: Multi-Object Tracking in Uniform Appearance and Diverse MotionPeize Sun, Jinkun Cao, Yi Jiang, Zehuan Yuan et al.CVPR 2022 · 305 citations
- MeMOT: Multi-Object Tracking with MemoryJiarui Cai, Mingze Xu, Wei Li, Yuanjun Xiong et al.CVPR 2022 · 216 citations
Related papers
- APPTracker: Improving Tracking Multiple Objects in Low-Frame-Rate VideosTao Zhou, Wenhan Luo, Zhiguo Shi, Jiming Chen et al.ACM MM 2022 · 10 citations
- CO-MOT: Boosting End-to-end Transformer-based Multi-Object Tracking via Coopetition Label Assignment and Shadow SetsFeng Yan, Weixin Luo, Yujie Zhong, Yiyang Gan et al.ICLR 2025
- MeMOTR: Long-Term Memory-Augmented Transformer for Multi-Object TrackingRuopeng Gao, Limin WangICCV 2023 · 143 citations
- From Detection to Association: Learning Discriminative Object Embeddings for Multi-Object TrackingYuqing Shao, Yuchen Yang, Rui Yu, Weilong Li et al.CVPR 2026 · 5 citations
- Standing Between Past and Future: Spatio-Temporal Modeling for Multi-Camera 3D Multi-Object TrackingZiqi Pang, Jie Li, Pavel Tokmakov, Dian Chen et al.CVPR 2023
