How to Train Your Deep Multi-Object Tracker
Yihong Xu, Aljosa Osep, Yutong Ban, Radu Horaud, Laura Leal-Taixé, Xavier Alameda-Pineda
摘要
The recent trend in vision-based multi-object tracking (MOT) is heading towards leveraging the representational power of deep learning to jointly learn to detect and track objects. However, existing methods train only certain submodules using loss functions that often do not correlate with established tracking evaluation measures such as Multi-Object Tracking Accuracy (MOTA) and Precision (MOTP). As these measures are not differentiable, the choice of appropriate loss functions for end-to-end training of multiobject tracking methods is still an open research problem. In this paper, we bridge this gap by proposing a differentiable proxy of MOTA and MOTP, which we combine in a loss function suitable for end-to-end training of deep multiobject trackers. As a key ingredient, we propose a Deep Hungarian Net (DHN) module that approximates the Hungarian matching algorithm. DHN allows estimating the correspondence between object tracks and ground truth objects to compute differentiable proxies of MOTA and MOTP, which are in turn used to optimize deep trackers directly. We experimentally demonstrate that the proposed differentiable framework improves the performance of existing multi-object trackers, and we establish a new state of the art on the MOTChallenge benchmark. Our code is publicly available from https://github.com/yihongXU/deepMOT .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper26
- MOTSynth: How Can Synthetic Data Help Pedestrian Detection and Tracking?Matteo Fabbri, Guillem Brasó, Gianluca Maugeri, Orcun Cetintas 等ICCV 2021 · 被引用 128 次
- Quo Vadis: Is Trajectory Forecasting the Key Towards Long-Term Multi-Object Tracking?Patrick Dendorfer, Vladimir Yugay, Aljosa Osep, Laura Leal-TaixéNeurIPS 2022 · 被引用 77 次
- Track without Appearance: Learn Box and Tracklet Embedding with Local and Global Motion Patterns for Vehicle TrackingGaoang Wang, Renshu Gu, Zuozhu Liu, Weijie Hu 等ICCV 2021 · 被引用 65 次
- Tracking People by Predicting 3D Appearance, Location and PoseJathushan Rajasegaran, Georgios Pavlakos, Angjoo Kanazawa, Jitendra MalikCVPR 2022 · 被引用 57 次
- A General Recurrent Tracking Framework without Real DataShuai Wang, Hao Sheng, Yang Zhang, Yubin Wu 等ICCV 2021 · 被引用 54 次
它引用的顶会 Paper2
相关 Paper
- Learnable Graph Matching: Incorporating Graph Partitioning With Deep Feature Learning for Multiple Object TrackingJiawei He, Zehao Huang, Naiyan Wang, Zhaoxiang ZhangCVPR 2021
- UniTrack: Differentiable Graph Representation Learning for Multi-Object TrackingBishoy Galoaa, Xiangyu Bai, Utsav Nandi, Sai Siddhartha Vivek Dhir Rangoju 等ICLR 2026 · 被引用 3 次
- Learning of Global Objective for Network Flow in Multi-Object TrackingShuai Li, Yu Kong, Hamid RezatofighiCVPR 2022 · 被引用 24 次
- A Unified Object Motion and Affinity Model for Online Multi-Object TrackingJunbo Yin, Wenguan Wang, Qinghao Meng, Ruigang Yang 等CVPR 2020
- Learning a Neural Solver for Multiple Object TrackingGuillem Brasó, Laura Leal-TaixéCVPR 2020
