Point-to-Region Loss for Semi-Supervised Point-Based Crowd Counting
Wei Lin, Chenyang Zhao, Antoni B. Chan
Abstract
Point detection has been developed to locate pedestrians in crowded scenes by training a counter through a pointto-point (P2P) supervision scheme. Despite its excellent localization and counting performance, training a pointbased counter still faces challenges concerning annotation labor: hundreds to thousands of points are required to annotate a single sample capturing a dense crowd. In this paper, we integrate point-based methods into a semisupervised counting framework based on pseudo-labeling, enabling the training of a counter with only a few annotated samples supplemented by a large volume of pseudolabeled data. However, during implementation, the training encounters issues as the confidence for pseudo-labels fails to be propagated to background pixels via the P2P. To tackle this challenge, we devise a point-specific activation map (PSAM) to visually interpret the phenomena occurring during the ill-posed training. Observations from the PSAM suggest that the feature map is excessively activated by the loss for unlabeled data, causing the decoder to misinterpret these over-activations as pedestrians. To mitigate this issue, we propose a point-to-region (P2R) scheme to substitute P2P, which segments out local regions rather than detects a point corresponding to a pedestrian for supervision. Consequently, pixels in the local region can share the same confidence with the corresponding pseudo points. Experimental results in both semi-supervised counting and unsupervised domain adaptation highlight the advantages of our method, illustrating P2R can resolve issues identified in PSAM. The code is available at https: //github.com/Elin24/P2RLoss .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext caa24688-8457-42a1-8d19-e8ab9dea85b5Cited by top-tier papers2
- Boosting Quantitive and Spatial Awareness for Zero-Shot Object CountingDa Zhang, Bingyu Li, Feiyu Wang, Zhiyuan Zhao et al.CVPR 2026 · 6 citations
- Decoupling What to Count and Where to See for Referring Expression CountingYuda Zou, Zijian Zhang, Yongchao XuAAAI 2026
Builds on28
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- Bayesian Loss for Crowd Count Estimation With Point SupervisionZhiheng Ma, Xing Wei, Xiaopeng Hong, Yihong GongICCV 2019 · 612 citations
- Rethinking Counting and Localization in Crowds: A Purely Point-Based FrameworkQingyu Song, Changan Wang, Zhengkai Jiang, Yabiao Wang et al.ICCV 2021 · 376 citations
- Boosting Crowd Counting via Multifaceted AttentionHui Lin, Zhiheng Ma, Rongrong Ji, Yaowei Wang et al.CVPR 2022 · 229 citations
- Localization in the Crowd with Topological ConstraintsShahira Abousamra, Minh Hoai, Dimitris Samaras, Chao ChenAAAI 2021 · 160 citations
Related papers
- Optimal Transport Minimization: Crowd Localization on Density Maps for Semi-Supervised CountingWei Lin, Antoni B. ChanCVPR 2023
- Weakly-Supervised Salient Object Detection Using Point SupervisonShuyong Gao, Wei Zhang, Yan Wang, Qianyu Guo et al.AAAI 2022 · 77 citations
- Weakly Supervised Semantic Segmentation via Progressive Confidence Region ExpansionXiangfeng Xu, Pinyi Zhang, Wenxuan Huang, Yunhang Shen et al.CVPR 2025
- Self-Supervised Object Localization with Joint Graph PartitionYukun Su, Guosheng Lin, Yun Hao, Yiwen Cao et al.AAAI 2022 · 17 citations
- Leveraging Self-Supervision for Cross-Domain Crowd CountingWeizhe Liu, Nikita Durasov, Pascal FuaCVPR 2022 · 43 citations
