Cartoon Face Recognition: A Benchmark Dataset
Yi Zheng, Yifan Zhao, Mengyuan Ren, He Yan, Xiangju Lu, Junhui Liu, Jia Li
Abstract
Recent years have witnessed increasing attention in cartoon media, powered by the strong demands of industrial applications. As the first step to understand this media, cartoon face recognition is a crucial but less-explored task with few datasets proposed. In this work, we first present a new challenging benchmark dataset, consisting of 389,678 images of 5,013 cartoon characters annotated with identity, bounding box, pose, and other auxiliary attributes. The dataset, named iCartoonFace, is currently the largest-scale, high-quality, rich-annotated, and spanning multiple occurrences in the field of image recognition, including near-duplications, occlusions, and appearance changes. In addition, we provide two types of annotations for cartoon media, i.e., face recognition, and face detection, with the help of a semi-automatic labeling algorithm. To further investigate this challenging dataset, we propose a multi-task domain adaptation approach that jointly utilizes the human and cartoon domain knowledge with three discriminative regularizations. We hence perform a benchmark analysis of the proposed dataset and verify the superiority of the proposed approach in the cartoon face recognition task. The dataset is available at https://iqiyi.cn/icartoonface.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9573190b-fe52-4e39-a2bd-8df6978ea78dCited by top-tier papers7
- Energy-based Hopfield Boosting for Out-of-Distribution DetectionClaus Hofmann, Simon Schmid, Bernhard Lehner, Daniel Klotz et al.NeurIPS 2024 · 19 citations
- Zero-Shot Character Identification and Speaker Prediction in Comics via Iterative Multimodal FusionYingxuan Li, Ryota Hinami, Kiyoharu Aizawa, Yusuke MatsuiACM MM 2024 · 3 citations
- NFT1000: A Cross-Modal Dataset For Non-Fungible Token RetrievalShuxun Wang, Yunfei Lei, Ziqi Zhang, Wei Liu et al.ACM MM 2024 · 3 citations
- Stylized-Face: A Million-Level Stylized Face Dataset for Face RecognitionZhengyuan Peng, Jianqing Xu, Yuge Huang, Jinkun Hao et al.ICCV 2025 · 1 citation
- Illuminating Visual Identity in Universal Multimodal EmbeddingsJiawei Cao, Junyi Feng, Jiashen Hua, Ziheng Huang et al.CVPR 2026 · 1 citation
Builds on1
Related papers
- StylizedFacePoint: Facial Landmark Detection for Stylized CharactersShengran Cheng, Chuhang Ma, Ye PanACM MM 2024 · 1 citation
- CADQ: Attribute-Consistent Face Cartoonization with Cross-modal Aligned and Deformable QuantizationYongjie Hu, Yifan Jiang, Ziyun Li, Fei Gao et al.ACM MM 2025
- Visual News: Benchmark and Challenges in News Image CaptioningFuxiao Liu, Yinghan Wang, Tianlu Wang, Vicente OrdonezEMNLP 2021 · 67 citations
- MagicCartoon: 3D Pose and Shape Estimation for Bipedal Cartoon CharactersYu-Pei Song, Yuantong Liu, Xiao Wu, Qi He et al.ACM MM 2024
- AnyTalk: Multi-modal Driven Multi-domain Talking Head GenerationYu Wang, Yunfei Liu, Fa-Ting Hong, Meng Cao et al.AAAI 2025 · 2 citations
