Focus On Details: Online Multi-Object Tracking with Diverse Fine-Grained Representation
Hao Ren, Shoudong Han, Huilin Ding, Ziwen Zhang, Hongwei Wang, Faquan Wang
摘要
Discriminative representation is essential to keep a unique identifier for each target in Multiple object tracking (MOT). Some recent MOT methods extract features of the bounding box region or the center point as identity embeddings. However, when targets are occluded, these coarsegrained global representations become unreliable. To this end, we propose exploring diverse fine-grained representation, which describes appearance comprehensively from global and local perspectives. This fine-grained representation requires high feature resolution and precise semantic information. To effectively alleviate the semantic misalignment caused by indiscriminate contextual information aggregation, Flow Alignment FPN (FAFPN) is proposed for multi-scale feature alignment aggregation. It generates semantic flow among feature maps from different resolutions to transform their pixel positions. Furthermore, we present a Multi-head Part Mask Generator (MPMG) to extract finegrained representation based on the aligned feature maps. Multiple parallel branches of MPMG allow it to focus on different parts of targets to generate local masks without label supervision. The diverse details in target masks facilitate fine-grained representation. Eventually, benefiting from a Shuffle-Group Sampling (SGS) training strategy with positive and negative samples balanced, we achieve stateof-the-art performance on MOT17 and MOT20 test sets. Even on DanceTrack, where the appearance of targets is extremely similar, our method significantly outperforms Byte-Track by 5.0% on HOTA and 5.6% on IDF1. Extensive experiments have proved that diverse fine-grained representation makes Re-ID great again in MOT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Hybrid-SORT: Weak Cues Matter for Online Multi-Object TrackingMingzhan Yang, Guangxin Han, Bin Yan, Wenhua Zhang 等AAAI 2024 · 被引用 171 次
- Towards Generalizable Multi-Object TrackingZheng Qin, Le Wang, Sanping Zhou, Panpan Fu 等CVPR 2024 · 被引用 21 次
- SAM2MOT: A Novel Paradigm of Multi-Object Tracking by SegmentationJunjie Jiang, Zelin Wang, Manqi Zhao, Yin Li 等AAAI 2026 · 被引用 19 次
- Self-Supervised Multi-Object Tracking with Path ConsistencyZijia Lu, Bing Shuai, Yanbei Chen, Zhenlin Xu 等CVPR 2024 · 被引用 13 次
- DeNoising-MOT: Towards Multiple Object Tracking with Severe OcclusionsTeng Fu, Xiaocong Wang, Haiyang Yu, Ke Niu 等ACM MM 2023 · 被引用 11 次
它引用的顶会 Paper12
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 被引用 6,042 次
- TransReID: Transformer-based Object Re-IdentificationShuting He, Hao Luo, Pichao Wang, Fan Wang 等ICCV 2021 · 被引用 1,172 次
- Tracking Without Bells and WhistlesPhilipp Bergmann, Tim Meinhardt, Laura Leal-TaixéICCV 2019 · 被引用 1,030 次
- DanceTrack: Multi-Object Tracking in Uniform Appearance and Diverse MotionPeize Sun, Jinkun Cao, Yi Jiang, Zehuan Yuan 等CVPR 2022 · 被引用 305 次
- Learning to Track with Object PermanencePavel Tokmakov, Jie Li, Wolfram Burgard, Adrien GaidonICCV 2021 · 被引用 241 次
相关 Paper
- From Detection to Association: Learning Discriminative Object Embeddings for Multi-Object TrackingYuqing Shao, Yuchen Yang, Rui Yu, Weilong Li 等CVPR 2026 · 被引用 5 次
- Focusing on Tracks for Online Multi-Object TrackingKyujin Shim, Kangwook Ko, Yujin Yang, Changick KimCVPR 2025
- MotionTrack: Learning Robust Short-Term and Long-Term Motions for Multi-Object TrackingZheng Qin, Sanping Zhou, Le Wang, Jinghai Duan 等CVPR 2023
- Improving Multiple Pedestrian Tracking by Track Management and Occlusion HandlingDaniel Stadler, Jürgen BeyererCVPR 2021
- DiffusionTrack: Diffusion Model for Multi-Object TrackingRun Luo, Zikai Song, Lintao Ma, Jinlin Wei 等AAAI 2024 · 被引用 77 次
