CPNet: Cartoon Parsing with Pixel and Part Correlation
Jian-Jun Qiao, Jie Zhang, Xiao Wu, Yu-Pei Song, Wei Li
Abstract
Cartoon parsing, the task of segmenting constituent parts such as heads, arms, and legs of cartoon characters, holds substantial significance for applications in the animation industry and emerging metaverse. Nonetheless, this domain presents considerable challenges stemming from complex visual appearances, irregular structures, abstract drawing styles, among other factors. In this paper, a novel Cartoon Parsing Network (CPNet) is introduced to address these challenges. CPNet skillfully leverages the spatial and semantic correlations of pixels to discern intricate and visually akin appearances. Furthermore, it employs both local and global correlations of constituent parts to differentiate irregular and abstract body sections. Specifically, the pixels of the cartoon image are interconnected by capitalizing on the spatial and semantic correlations. To this end, a center point predictor, working in tandem with a pixel-aware attention, facilitates the exploration of pixel-level correlation learning. Additionally, the various constituent parts are meticulously organized to resonate with the intrinsic physiological structure of a cartoon character. The character's graph structure is assembled and analyzed by an edge-aware graph neural network, thereby linking adjacent parts and assimilating local correlations. A part-guided non-local attention mechanism is fashioned to correlate individual parts with the entire body, thereby modeling global connections. In addition, a new dataset named CartoonSet is curated and annotated explicitly for cartoon parsing. Experiments carried out on both cartoon parsing and human parsing datasets yield compelling results, thereby attesting to the efficacy and innovativeness of the proposed method.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 9eaafc55-9765-4efc-981f-e1db30528f69Cited by top-tier papers1
Ask how each one uses itRelated papers
- CartoonNet: Cartoon Parsing with Semantic Consistency and Structure CorrelationJian-Jun Qiao, Meng-Yu Duan, Xiao Wu, Yu-Pei SongACM MM 2024
- CAPNet: Cartoon Animal Parsing with Spatial Learning and Structural ModelingJian-Jun Qiao, Meng-Yu Duan, Xiao Wu, Wei LiACM MM 2024
- Part-Aware Context Network for Human ParsingXiaomei Zhang, Yingying Chen, Bingke Zhu, Jinqiao Wang et al.CVPR 2020
- Hierarchical Human Parsing With Typed Part-Relation ReasoningWenguan Wang, Hailong Zhu, Jifeng Dai, Yanwei Pang et al.CVPR 2020
- CDGNet: Class Distribution Guided Network for Human ParsingKunliang Liu, Ouk Choi, Jianming Wang, Wonjun HwangCVPR 2022 · 44 citations
