Towards Discriminative Representation: Multi-view Trajectory Contrastive Learning for Online Multi-object Tracking
En Yu, Zhuoling Li, Shoudong Han
Abstract
Discriminative representation is crucial for the association step in multi-object tracking. Recent work mainly utilizes features in single or neighboring frames for constructing metric loss and empowering networks to extract representation of targets. Although this strategy is effective, it fails to fully exploit the information contained in a whole trajectory. To this end, we propose a strategy, namely multi-view trajectory contrastive learning, in which each trajectory is represented as a center vector. By maintaining all the vectors in a dynamically updated memory bank, a trajectory-level contrastive loss is devised to explore the inter-frame information in the whole trajectories. Besides, in this strategy, each target is represented as multiple adaptively selected keypoints rather than a pre-defined anchor or center. This design allows the network to generate richer representation from multiple views of the same target, which can better characterize occluded objects. Additionally, in the inference stage, a similarity-guided feature fusion strategy is developed for further boosting the quality of the trajectory representation. Extensive experiments have been conducted on MOTChallenge to verify the effectiveness of the proposed techniques. The experimental results indicate that our method has surpassed preceding trackers and established new state-of-the-art performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 76b6848a-5adf-457c-9007-8cdb59f62897Cited by top-tier papers18
- CTVIS: Consistent Training for Online Video Instance SegmentationKaining Ying, Qing Zhong, Weian Mao, Zhenhua Wang et al.ICCV 2023 · 72 citations
- Neural Collapse with Normalized Features: A Geometric Analysis over the Riemannian ManifoldCan Yaras, Peng Wang, Zhihui Zhu, Laura Balzano et al.NeurIPS 2022 · 60 citations
- Spherical Space Feature Decomposition for Guided Depth Map Super-ResolutionZixiang Zhao, Jiangshe Zhang, Xiang Gu, Chengli Tan et al.ICCV 2023 · 55 citations
- Generalizing Multiple Object Tracking to Unseen Domains by Introducing Natural Language RepresentationEn Yu, Songtao Liu, Zhuoling Li, Jinrong Yang et al.AAAI 2023 · 22 citations
- Uncertainty-aware Unsupervised Multi-Object TrackingKai Liu, Sheng Jin, Zhihang Fu, Ze Chen et al.ICCV 2023 · 21 citations
Builds on15
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- Tracking Without Bells and WhistlesPhilipp Bergmann, Tim Meinhardt, Laura Leal-TaixéICCV 2019 · 1,030 citations
Related papers
- From Detection to Association: Learning Discriminative Object Embeddings for Multi-Object TrackingYuqing Shao, Yuchen Yang, Rui Yu, Weilong Li et al.CVPR 2026 · 5 citations
- DOVTrack: Data-Efficient Open-Vocabulary TrackingZekun Qian, Ruize Han, Zhixiang Wang, Junhui Hou et al.NeurIPS 2025 · 1 citation
- Tracking without Label: Unsupervised Multiple Object Tracking via Contrastive Similarity LearningSha Meng, Dian Shao, Jiacheng Guo, Shan GaoICCV 2023 · 14 citations
- Multiple Object Tracking as ID PredictionRuopeng Gao, Ji Qi, Limin WangCVPR 2025
- Quasi-Dense Similarity Learning for Multiple Object TrackingJiangmiao Pang, Linlu Qiu, Xia Li, Haofeng Chen et al.CVPR 2021
