Motion Adaptive Pose Estimation from Compressed Videos
Zhipeng Fan, Jun Liu, Yao Wang
Abstract
Human pose estimation from videos has many real-world applications. Existing methods focus on applying models with a uniform computation profile on fully decoded frames, ignoring the freely-available motion signals and motion-compensation residuals from the compressed stream. A novel model, called Motion Adaptive Pose Net is proposed to exploit the compressed streams to efficiently decode pose sequences from videos. The model incorporates a Motion Compensated ConvLSTM to propagate the spatially aligned features, along with an adaptive gate to dynamically determine if the computationally expensive features should be extracted from fully decoded frames to compensate the motion-warped features, solely based on the residual errors. Leveraging the informative yet readily available signals from compressed streams, we propagate the latent features through our Motion Adaptive Pose Net efficiently Our model outperforms the state-of-the-art models in pose-estimation accuracy on two widely used datasets with only around half of the computation complexity.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e6f73fb1-9728-4e51-917e-78e8ba0e0fe0Cited by top-tier papers8
- Temporal Feature Alignment and Mutual Information Maximization for Video-Based Human Pose EstimationZhenguang Liu, Runyang Feng, Haoming Chen, Shuang Wu et al.CVPR 2022 · 76 citations
- MixSynthFormer: A Transformer Encoder-like Structure with Mixed Synthetic Self-attention for Efficient Human Pose EstimationYuran Sun, Alan William Dougherty, Zhuoying Zhang, Yi-King Choi et al.ICCV 2023 · 6 citations
- : Discrete Diffusion Model for Occluded 3D Human Pose EstimationWeiquan Wang, Jun Xiao, Chunping Wang, Wei Liu et al.NeurIPS 2024 · 4 citations
- High-Resolution Spatiotemporal Modeling with Global-Local State Space Models for Video-Based Human Pose EstimationRunyang Feng, Hyung Jin Chang, Tze Ho Elden Tse, Boeun Kim et al.ICCV 2025 · 2 citations
- Track-On: Transformer-based Online Point Tracking with MemoryGörkay Aydemir, Xiongyi Cai, Weidi Xie, Fatma GüneyICLR 2025
Builds on2
Related papers
- Structure-Preserving Motion Estimation for Learned Video CompressionHan Gao, Jinzhong Cui, Mao Ye, Shuai Li et al.ACM MM 2022 · 16 citations
- MVFlow: Deep Optical Flow Estimation of Compressed Videos with Motion Vector PriorShili Zhou, Xuhao Jiang, Weimin Tan, Ruian He et al.ACM MM 2023 · 8 citations
- MotionRefineNet: Fine-Grained Pose Sequence Smoothing and RefinementHaolun Li, Weihuang Liu, Jiateng Liu, Zhenhua Tang et al.ACM MM 2025 · 9 citations
- Deep Dual Consecutive Network for Human Pose EstimationZhenguang Liu, Haoming Chen, Runyang Feng, Shuang Wu et al.CVPR 2021
- Joint-Motion Mutual Learning for Pose Estimation in VideoSifan Wu, Haipeng Chen, Yifang Yin, Sihao Hu et al.ACM MM 2024 · 10 citations
