Lite Pose: Efficient Architecture Design for 2D Human Pose Estimation
Yihan Wang, Muyang Li, Han Cai, Wei-Ming Chen, Song Han
Abstract
Pose estimation plays a critical role in human-centered vision applications. However, it is difficult to deploy state-of-the-art HRNet-based pose estimation models on resource-constrained edge devices due to the high computational cost (more than 150 GMACs per frame). In this paper, we study efficient architecture design for real-time multi-person pose estimation on edge. We reveal that HRNet's high-resolution branches are redundant for models at the low-computation region via our gradual shrinking experiments. Removing them improves both efficiency and performance. Inspired by this finding, we design LitePose, an efficient single-branch architecture for pose estimation, and introduce two simple approaches to enhance the capacity of LitePose, including fusion deconv head and large kernel conv. On mobile platforms, LitePose reduces the latency by up to <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> without sacrificing performance, compared with prior state-of-the-art efficient pose estimation models, pushing the frontier of real-time multi-person pose estimation on edge. Our code and pretrained models are released at https://github.com/mit-han-lab/litepose.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- EfficientViT: Lightweight Multi-Scale Attention for High-Resolution Dense PredictionHan Cai, Junyan Li, Muyan Hu, Chuang Gan et al.ICCV 2023 · 265 citations
- Differentially Private 2D Human Pose EstimationKaushik Bhargav Sivangi, Paul Henderson, Fani DeligianniCVPR 2026 · 1 citation
- Sequential Joint Dependency Aware Human Pose Estimation with State Space ModelHanxi Yin, Shaodi You, Jungong Han, Zhixiang ChenAAAI 2025 · 1 citation
- Human Pose as Compositional TokensZigang Geng, Chunyu Wang, Yixuan Wei, Ze Liu et al.CVPR 2023
- JAWS: Just A Wild Shot for Cinematic Transfer in Neural Radiance FieldsXi Wang, Robin Courant, Jinglei Shi, Éric Marchand et al.CVPR 2023
Builds on8
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le et al.ICCV 2019 · 9,163 citations
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
- MetaPruning: Meta Learning for Automatic Neural Network Channel PruningZechun Liu, Haoyuan Mu, Xiangyu Zhang, Zichao Guo et al.ICCV 2019 · 633 citations
- Ansor: Generating High-Performance Tensor Programs for Deep LearningLianmin Zheng, Chengfan Jia, Minmin Sun, Zhao Wu et al.OSDI 2020 · 551 citations
- Lite Transformer with Long-Short Range AttentionZhanghao Wu, Zhijian Liu, Ji Lin, Yujun Lin et al.ICLR 2020 · 379 citations
Related papers
- Lite-HRNet: A Lightweight High-Resolution NetworkChangqian Yu, Bin Xiao, Changxin Gao, Lu Yuan et al.CVPR 2021
- DynPose: Largely Improving the Efficiency of Human Pose Estimation by a Simple Dynamic FrameworkYalong Xu, Lin Zhao, Chen Gong, Guangyu Li et al.CVPR 2025
- Single-Network Whole-Body Pose EstimationGines Hidalgo Martinez, Yaadhav Raaj, Haroon Idrees, Donglai Xiang et al.ICCV 2019 · 115 citations
- SDPose: Tokenized Pose Estimation via Circulation-Guide Self-DistillationSichen Chen, Yingyi Zhang, Siming Huang, Ran Yi et al.CVPR 2024
- AdaptivePose: Human Parts as Adaptive PointsYabo Xiao, Xiaojuan Wang, Dongdong Yu, Guoli Wang et al.AAAI 2022 · 25 citations
