Spatiotemporal Feature Residual Propagation for Action Prediction
He Zhao, Rick Wildes
摘要
Recognizing actions from limited preliminary video observations has seen considerable recent progress. Typically, however, such progress has been had without explicitly modeling fine-grained motion evolution as a potentially valuable information source. In this study, we address this task by investigating how action patterns evolve over time in a spatial feature space. There are three key components to our system. First, we work with intermediate-layer ConvNet features, which allow for abstraction from raw data, while retaining spatial layout, which is sacrificed in approaches that rely on vectorized global representations. Second, instead of propagating features per se, we propagate their residuals across time, which allows for a compact representation that reduces redundancy while retaining essential information about evolution over time. Third, we employ a Kalman filter to combat error build-up and unify across prediction start times. Extensive experimental results on the JHMDB21, UCF101 and BIT datasets show that our approach leads to a new state-of-the-art in action prediction.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Bifold and Semantic Reasoning for Pedestrian Behavior PredictionAmir Rasouli, Mohsen Rohani, Jun LuoICCV 2021 · 被引用 69 次
- Distillation Using Oracle Queries for Transformer-based Human-Object Interaction DetectionXian Qu, Changxing Ding, Xingao Li, Xubin Zhong 等CVPR 2022 · 被引用 48 次
- Anticipating Future Relations via Graph Growing for Action PredictionXinxiao Wu, Jianwei Zhao, Ruiqi WangAAAI 2021 · 被引用 24 次
- EAST: Early Action Prediction Sampling Strategy with Token MaskingIva Sović, Ivan Martinović, Marin OršićICLR 2026
- Anticipating Human Actions by Correlating Past With the Future With Jaccard Similarity MeasuresBasura Fernando, Samitha HerathCVPR 2021
相关 Paper
- STM: SpatioTemporal and Motion Encoding for Action RecognitionBoyuan Jiang, Mengmeng Wang, Weihao Gan, Wei Wu 等ICCV 2019 · 被引用 442 次
- TEA: Temporal Excitation and Aggregation for Action RecognitionYan Li, Bin Ji, Xintian Shi, Jianguo Zhang 等CVPR 2020
- The Wisdom of Crowds: Temporal Progressive Attention for Early Action PredictionAlexandros Stergiou, Dima DamenCVPR 2023
- Evolving Space-Time Neural Architectures for VideosA. J. Piergiovanni, Anelia Angelova, Alexander Toshev, Michael S. RyooICCV 2019 · 被引用 62 次
- V4D: 4D Convolutional Neural Networks for Video-level Representation LearningShiwen Zhang, Sheng Guo, Weilin Huang, Matthew R. Scott 等ICLR 2020 · 被引用 81 次
