Video Face Clustering With Unknown Number of Clusters
Makarand Tapaswi, Marc T. Law, Sanja Fidler
摘要
Understanding videos such as TV series and movies requires analyzing who the characters are and what they are doing. We address the challenging problem of clustering face tracks based on their identity. Different from previous work in this area, we choose to operate in a realistic and difficult setting where: (i) the number of characters is not known a priori; and (ii) face tracks belonging to minor or background characters are not discarded. To this end, we propose Ball Cluster Learning (BCL), a supervised approach to carve the embedding space into balls of equal size, one for each cluster. The learned ball radius is easily translated to a stopping criterion for iterative merging algorithms. This gives BCL the ability to estimate the number of clusters as well as their assignment, achieving promising results on commonly used datasets. We also present a thorough discussion of how existing metric learning literature can be adapted for this task.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Deep Open Intent Classification with Adaptive Decision BoundaryHanlei Zhang, Hua Xu, Ting-En LinAAAI 2021 · 被引用 127 次
- A Theoretical Analysis of the Number of Shots in Few-Shot LearningTianshi Cao, Marc T. Law, Sanja FidlerICLR 2020 · 被引用 75 次
- DeepDPM: Deep Clustering With an Unknown Number of ClustersMeitar Ronen, Shahaf E. Finder, Oren FreifeldCVPR 2022 · 被引用 66 次
- AutoAD II: The Sequel - Who, When, and What in Movie Audio DescriptionTengda Han, Max Bain, Arsha Nagrani, Gül Varol 等ICCV 2023 · 被引用 55 次
- AVA-AVD: Audio-visual Speaker Diarization in the WildEric Zhongcong Xu, Zeyang Song, Satoshi Tsutsui, Chao Feng 等ACM MM 2022 · 被引用 34 次
相关 Paper
- Robust Actor Recognition in Entertainment Multimedia at ScaleAbhinav Aggarwal, Yash Pandya, Lokesh A. Ravindranathan, Laxmi S. Ahire 等ACM MM 2022 · 被引用 4 次
- Learned Trajectory Embedding for Subspace ClusteringYaroslava Lochman, Carl Olsson, Christopher ZachCVPR 2024 · 被引用 5 次
- Tracklet Self-Supervised Learning for Unsupervised Person Re-IdentificationGuile Wu, Xiatian Zhu, Shaogang GongAAAI 2020 · 被引用 97 次
- CLIP-Cluster: CLIP-Guided Attribute Hallucination for Face ClusteringShuai Shen, Wanhua Li, Xiaobing Wang, Dafeng Zhang 等ICCV 2023 · 被引用 19 次
- VideoCutLER: Surprisingly Simple Unsupervised Video Instance SegmentationXudong Wang, Ishan Misra, Ziyun Zeng, Rohit Girdhar 等CVPR 2024
