Weakly Supervised Video Individual Counting
Xinyan Liu, Guorong Li, Yuankai Qi, Ziheng Yan, Zhenjun Han, Anton van den Hengel, Ming-Hsuan Yang, Qingming Huang
摘要
Video Individual Counting (VIC) aims to predict the number of unique individuals in a single video. Existing methods learn representations based on trajectory labels for individuals, which are annotation-expensive. To provide a more realistic reflection of the underlying practical challenge, we introduce a weakly supervised VIC task, wherein trajectory labels are not provided. Instead, two types of labels are provided to indicate traffic entering the field of view (inflow) and leaving the field view (outflow). We also propose the first solution as a baseline that formulates the task as a weakly supervised contrastive learning problem under group-level matching. In doing so, we devise an end-to-end trainable soft contrastive loss to drive the network to distinguish inflow, outflow, and the remaining. To facilitate future study in this direction, we generate annotations from the existing VIC datasets SenseCrowd and CroHD and also build a new dataset, UAVVIC. Extensive results show that our baseline weakly supervised method outperforms supervised methods, and thus, little information is lost in the transition to the more practically relevant weakly supervised task. The code and trained model can be found at CGNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper6
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- SMILEtrack: SiMIlarity LEarning for Occlusion-Aware Multiple Object TrackingYu-Hsiang Wang, Jun-Wei Hsieh, Ping-Yang Chen, Ming-Ching Chang 等AAAI 2024 · 被引用 96 次
- CWCL: Cross-Modal Transfer with Continuously Weighted Contrastive LossRakshith Sharma Srinivasa, Jaejin Cho, Chouchang Yang, Yashas Malur Saidutta 等NeurIPS 2023 · 被引用 25 次
- Discriminative Appearance Modeling With Multi-Track Pooling for Real-Time Multi-Object TrackingChanho Kim, Fuxin Li, Mazen Alotaibi, James M. RehgCVPR 2021
- Tracking Pedestrian Heads in Dense CrowdRamana Sundararaman, Cedric De Almeida Braga, Éric Marchand, Julien PettréCVPR 2021
相关 Paper
- Flowing Crowd to Count Flows: A Self-Supervised Framework for Video Individual CountingFeng-Kai Huang, Bo-Lun Huang, Li-Wu Tsao, Jhih-Ciang Wu 等ACM MM 2025 · 被引用 1 次
- Video Individual Counting for Moving DronesYaowu Fan, Jia Wan, Tao Han, Antoni B. Chan 等ICCV 2025 · 被引用 1 次
- DR.VIC: Decomposition and Reasoning for Video Individual CountingTao Han, Lei Bai, Junyu Gao, Qi Wang 等CVPR 2022 · 被引用 18 次
- Prototype-Guided Dual-Transformer Reasoning for Video Individual CountingRui Li, Yishu Liu, Huafeng Li, Jinxing Li 等ACM MM 2024 · 被引用 2 次
- TLMA: Mitigating the Impact of Weakly Labeled Information for Video Anomaly DetectionRong Xu, Runqi Wang, Yingjun Zhang, Tao Tao 等CVPR 2026
