Localization with Sampling-Argmax
Jiefeng Li, Tong Chen, Ruiqi Shi, Yujing Lou, Yong-Lu Li, Cewu Lu
摘要
Soft-argmax operation is commonly adopted in detection-based methods to localize the target position in a differentiable manner. However, training the neural network with soft-argmax makes the shape of the probability map unconstrained. Consequently, the model lacks pixel-wise supervision through the map during training, leading to performance degradation. In this work, we propose sampling-argmax, a differentiable training method that imposes implicit constraints to the shape of the probability map by minimizing the expectation of the localization error. To approximate the expectation, we introduce a continuous formulation of the output distribution and develop a differentiable sampling process. The expectation can be approximated by calculating the average error of all samples drawn from the output distribution. We show that sampling-argmax can seamlessly replace the conventional soft-argmax operation on various localization tasks. Comprehensive experiments demonstrate the effectiveness and flexibility of the proposed method. Code is available at https://github.com/Jeff-sjtu/sampling-argmax .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- LocLLM: Exploiting Generalizable Human Keypoint Localization via Large Language ModelDongkai Wang, Shiyu Xuan, Shiliang ZhangCVPR 2024 · 被引用 15 次
- Spatial-Aware Regression for Keypoint LocalizationDongkai Wang, Shiliang ZhangCVPR 2024 · 被引用 5 次
- TIC-TAC: A Framework For Improved Covariance Estimation In Deep Heteroscedastic RegressionMegh Shukla, Mathieu Salzmann, Alexandre AlahiICML 2024 · 被引用 5 次
- Towards Self-Supervised Covariance Estimation in Deep Heteroscedastic RegressionMegh Shukla, Aziz Shameem, Mathieu Salzmann, Alexandre AlahiICLR 2025
- Generalizable Object Keypoint Localization from Generative PriorsDongkai Wang, Jiang Duan, Liangjian Wen, Shiyu Xuan 等CVPR 2025
它引用的顶会 Paper9
- Camera Distance-Aware Top-Down Approach for 3D Multi-Person Pose Estimation From a Single RGB ImageGyeongsik Moon, Ju Yong Chang, Kyoung Mu LeeICCV 2019 · 被引用 368 次
- DeepPruner: Learning Efficient Stereo Matching via Differentiable PatchMatchShivam Duggal, Shenlong Wang, Wei-Chiu Ma, Rui Hu 等ICCV 2019 · 被引用 300 次
- Human Pose Regression with Residual Log-likelihood EstimationJiefeng Li, Siyuan Bian, Ailing Zeng, Can Wang 等ICCV 2021 · 被引用 286 次
- Learning to Orient Surfaces by Self-supervised Spherical CNNsRiccardo Spezialetti, Federico Stella, Marlon Marcon, Luciano Silva 等NeurIPS 2020 · 被引用 48 次
- Attention-Driven Cropping for Very High Resolution Facial Landmark DetectionPrashanth Chandran, Derek Bradley, Markus Gross, Thabo BeelerCVPR 2020
相关 Paper
- Dive Deeper Into Integral Pose RegressionKerui Gu, Linlin Yang, Angela YaoICLR 2022 · 被引用 17 次
- Heatmap Regression without Soft-Argmax for Facial Landmark DetectionChiao-An Yang, Raymond A. YehICCV 2025 · 被引用 3 次
- Unsupervised Object Detection with Theoretical GuaranteesMarian Longa, João F. HenriquesNeurIPS 2024 · 被引用 1 次
- Dual-Gradients Localization Framework for Weakly Supervised Object LocalizationChuangchuang Tan, Guanghua Gu, Tao Ruan, Shikui Wei 等ACM MM 2020 · 被引用 19 次
- GrooMeD-NMS: Grouped Mathematically Differentiable NMS for Monocular 3D Object DetectionAbhinav Kumar, Garrick Brazil, Xiaoming LiuCVPR 2021
