Tracking through Severe Occlusion via Event-Derived Transient Cues
Hao Dong, Yujin Liu, Haoyue Liu, Zhenyu Wang, Shihan Peng, Zhiwei Shi, Yi Chang, Luxin Yan
Abstract
Tracking targets with high-speed and nonlinear motion under occlusion remains challenging due to spatial appearance deprivation and temporal trajectory fragmentation caused by missing visual cues. Existing methods typically either dynamically update templates to maintain appearance similarity or employ autoregressive models to predict targets from historical trajectories. However, these methods are ineffective under severe occlusion owing to template contamination and limited frame rates for complex motion. In this work, we observe that occlusion inherently degrades the spatial matching mechanism, highlighting the importance of temporal cues. Meanwhile, event cameras with microsecond-level temporal resolution provide transient dynamic cues that facilitate modeling nonlinear motion. In light of this, we propose EvoTrack, an occlusion-robust tracking framework via event-derived transient evolution, which comprises event-based motion autoregression and target-aware appearance matching. Specifically, for motion autoregression, the fine-grained timestamps of events naturally encode the target's direction and speed, motivating a bidirectional motion consistency that constrains interframe displacement prediction under nonlinear motion. For appearance matching, we adopt a Gaussian masking strategy to simulate occlusion degradation, guiding the model to focus on target regions and learn invariant representations. Furthermore, we build a pixel-aligned Frame-Event tracking dataset with higher spatial resolution and explicit occlusion labels. Extensive experiments demonstrate the effectiveness of EvoTrack in challenging occlusion scenes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 36ca89eb-afd8-452e-956d-f8cd0e90b132Builds on35
- MixFormer: End-to-End Tracking with Iterative Mixed AttentionYutao Cui, Cheng Jiang, Limin Wang, Gangshan WuCVPR 2022 · 746 citations
- Spiking Transformers for Event-based Single Object TrackingJiqing Zhang, Bo Dong, Haiwei Zhang, Jianchuan Ding et al.CVPR 2022 · 171 citations
- Object Tracking by Jointly Exploiting Frame and Event DomainJiqing Zhang, Xin Yang, Yingkai Fu, Xiaopeng Wei et al.ICCV 2021 · 141 citations
- SMILEtrack: SiMIlarity LEarning for Occlusion-Aware Multiple Object TrackingYu-Hsiang Wang, Jun-Wei Hsieh, Ping-Yang Chen, Ming-Ching Chang et al.AAAI 2024 · 96 citations
- Single-Model and Any-Modality for Video Object TrackingZongwei Wu, Jilai Zheng, Xiangxuan Ren, Florin-Alexandru Vasluianu et al.CVPR 2024 · 78 citations
Related papers
- MATE: Motion-Augmented Temporal Consistency for Event-Based Point TrackingHan Han, Wei Zhai, Yang Cao, Bin Li et al.ICCV 2025 · 3 citations
- Event6D: Event-based Novel Object 6D Pose TrackingJae-Young Kang, Hoonhee Cho, Taeyeop Lee, Minjun Kang et al.CVPR 2026 · 4 citations
- TimeTracker: Event-based Continuous Point Tracking for Video Frame Interpolation with Non-linear MotionHaoyue Liu, Jinghan Xu, Yi Chang, Hanyu Zhou et al.CVPR 2025
- Learning Visual Motion Segmentation Using Event SurfacesAnton Mitrokhin, Zhiyuan Hua, Cornelia Fermüller, Yiannis AloimonosCVPR 2020
- Event-Based Motion Deblurring Using Task-Oriented 3D Gaussian Event RepresentationsShengdong Xue, Haoxiang Ma, Hao Chen, Zhen Yang et al.CVPR 2026
