Multiple Planar Object Tracking
Zhicheng Zhang, Shengzhe Liu, Jufeng Yang
摘要
Tracking both location and pose of multiple planar objects (MPOT) is of great significance to numerous real-world applications. The greater degree-of-freedom of planar objects compared with common objects makes MPOT far more challenging than well-studied object tracking, especially when occlusion occurs. To address this challenging task, we are inspired by amodal perception that humans jointly track visible and invisible parts of the target, and propose a tracking framework that unifies appearance perception and occlusion reasoning. Specifically, we present a dual-branch network to track the visible part of planar objects, including vertexes and mask. Then, we develop an occlusion area localization strategy to infer the invisible part, i.e., the occluded region, followed by a two-stream attention network finally refining the prediction. To alleviate the lack of data in this field, we build the first large-scale benchmark dataset, namely MPOT-3K. It consists of 3,717 planar objects from 356 videos and contains 148,896 frames together with 687,417 annotations. The collected planar objects have 9 motion patterns and the videos are shot in 6 types of indoor and outdoor scenes. Extensive experiments demonstrate the superiority of our proposed method on the newly developed MPOT-3K as well as other two popular single planar object tracking datasets. The code and MPOT-3K dataset are released on https://zzcheng.top/MPOT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Adapt or Perish: Adaptive Sparse Transformer with Attentive Feature Refinement for Image RestorationShihao Zhou, Duosheng Chen, Jinshan Pan, Jinglei Shi 等CVPR 2024 · 被引用 137 次
- ExtDM: Distribution Extrapolation Diffusion Model for Video PredictionZhicheng Zhang, Junyao Hu, Wentao Cheng, Danda Pani Paudel 等CVPR 2024 · 被引用 24 次
- MCNet: Rethinking the Core Ingredients for Accurate and Efficient Homography EstimationHaokai Zhu, Si-Yuan Cao, Jianxin Hu, Sitong Zuo 等CVPR 2024 · 被引用 18 次
- LAKE-RED: Camouflaged Images Generation by Latent Background Knowledge Retrieval-Augmented DiffusionPancheng Zhao, Peng Xu, Pengda Qin, Deng-Ping Fan 等CVPR 2024 · 被引用 14 次
- MART: Masked Affective RepresenTation Learning via Masked Temporal Distribution DistillationZhicheng Zhang, Pancheng Zhao, Eunil Park, Jufeng YangCVPR 2024 · 被引用 11 次
它引用的顶会 Paper17
- Video Object Segmentation Using Space-Time Memory NetworksSeoung Wug Oh, Joon-Young Lee, Ning Xu, Seon Joo KimICCV 2019 · 被引用 845 次
- Stratified Transformer for 3D Point Cloud SegmentationXin Lai, Jianhui Liu, Li Jiang, Liwei Wang 等CVPR 2022 · 被引用 494 次
- Video Instance Segmentation with a Propose-Reduce ParadigmHuaijia Lin, Ruizheng Wu, Shu Liu, Jiangbo Lu 等ICCV 2021 · 被引用 110 次
- Spatial Pruned Sparse Convolution for Efficient 3D Object DetectionJianhui Liu, Yukang Chen, Xiaoqing Ye, Zhuotao Tian 等NeurIPS 2022 · 被引用 57 次
- Planar Surface Reconstruction from Sparse ViewsLinyi Jin, Shengyi Qian, Andrew Owens, David F. FouheyICCV 2021 · 被引用 51 次
相关 Paper
- Human De-Occlusion: Invisible Perception and Recovery for HumansQiang Zhou, Shiyin Wang, Yitong Wang, Zilong Huang 等CVPR 2021
- PlanarTrack: A Large-scale Challenging Benchmark for Planar Object TrackingXinran Liu, Xiaoqiong Liu, Ziruo Yi, Xin Zhou 等ICCV 2023 · 被引用 2 次
- Amodal Panoptic SegmentationRohit Mohan, Abhinav ValadaCVPR 2022 · 被引用 49 次
- PoseTrack21: A Dataset for Person Search, Multi-Object Tracking and Multi-Person Pose TrackingAndreas Doering, Di Chen, Shanshan Zhang, Bernt Schiele 等CVPR 2022 · 被引用 47 次
- Learning to Track with Object PermanencePavel Tokmakov, Jie Li, Wolfram Burgard, Adrien GaidonICCV 2021 · 被引用 241 次
