The Center of Attention: Center-Keypoint Grouping via Attention for Multi-Person Pose Estimation
Guillem Brasó, Nikita Kister, Laura Leal-Taixé
摘要
We introduce CenterGroup, an attention-based framework to estimate human poses from a set of identity-agnostic keypoints and person center predictions in an image. Our approach uses a transformer to obtain context-aware embeddings for all detected keypoints and centers and then applies multi-head attention to directly group joints into their corresponding person centers. While most bottom-up methods rely on non-learnable clustering at inference, CenterGroup uses a fully differentiable attention mechanism that we train end-to-end together with our keypoint detector. As a result, our method obtains state-of-the-art performance with up to 2.5x faster inference time than competing bottom-up approaches. Our code is available at https://github.com/dvl-tum/center-group
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Contextual Instance Decoupling for Robust Multi-Person Pose EstimationDongkai Wang, Shiliang ZhangCVPR 2022 · 被引用 73 次
- QueryPose: Sparse Multi-Person Pose Regression via Spatial-Aware Part-Level QueryYabo Xiao, Kai Su, Xiaojuan Wang, Dongdong Yu 等NeurIPS 2022 · 被引用 32 次
- Rethinking pose estimation in crowds: overcoming the detection information bottleneck and ambiguityMu Zhou, Lucas Stoffl, Mackenzie Weygandt Mathis, Alexander MathisICCV 2023 · 被引用 28 次
- Mutual Adaptive Reasoning for Monocular 3D Multi-Person Pose EstimationJuze Zhang, Jingya Wang, Ye Shi, Fei Gao 等ACM MM 2022 · 被引用 15 次
- Keypoint-Augmented Self-Supervised Learning for Medical Image Segmentation with Limited AnnotationZhangsihao Yang, Mengwei Ren, Kaize Ding, Guido Gerig 等NeurIPS 2023 · 被引用 12 次
它引用的顶会 Paper14
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- Generative Pretraining From PixelsMark Chen, Alec Radford, Rewon Child, Jeffrey Wu 等ICML 2020 · 被引用 1,773 次
- Everybody Dance NowCaroline Chan, Shiry Ginosar, Tinghui Zhou, Alexei A. EfrosICCV 2019 · 被引用 840 次
- Single-Stage Multi-Person Pose MachinesXuecheng Nie, Jiashi Feng, Jianfeng Zhang, Shuicheng YanICCV 2019 · 被引用 246 次
相关 Paper
- Bottom-Up Human Pose Estimation via Disentangled Keypoint RegressionZigang Geng, Ke Sun, Bin Xiao, Zhaoxiang Zhang 等CVPR 2021
- Group Pose: A Simple Baseline for End-to-End Multi-person Pose EstimationHuan Liu, Qiang Chen, Zichang Tan, Jiang-Jiang Liu 等ICCV 2023 · 被引用 50 次
- Keypoint CommunitiesDuncan Zauss, Sven Kreiss, Alexandre AlahiICCV 2021 · 被引用 19 次
- Optimizing Human Pose Estimation Through Focused Human and Joint RegionsYingying Jiao, Zhigang Wang, Zhenguang Liu, Shaojing Fan 等AAAI 2025 · 被引用 4 次
- Simple Pose: Rethinking and Improving a Bottom-up Approach for Multi-Person Pose EstimationJia Li, Wen Su, Zengfu WangAAAI 2020 · 被引用 104 次
