CAPNet: Cartoon Animal Parsing with Spatial Learning and Structural Modeling
Jian-Jun Qiao, Meng-Yu Duan, Xiao Wu, Wei Li
Abstract
Cartoon animal parsing aims to segment the body parts such as heads, arms, legs and tails of cartoon animals. Different from previous parsing tasks, cartoon animal parsing faces new challenges, including irregular body structures, abstract drawing styles and diverse animal categories. Existing methods have difficulties when addressing these challenges caused by the spatial and structural properties of cartoon animals. To address these challenges, a novel spatial learning and structural modeling network, named CAPNet, is proposed for cartoon animal parsing. It aims to address the critical problems of spatial perception, structure modeling and spatial-structural consistency learning. A spatial-aware learning module integrates deformable convolutions to learn spatial features of diverse cartoon animals. The multi-task edge and center point prediction mechanism is incorporated to capture the intricate spatial patterns. A structural modeling method is proposed to model the complex structural representations of cartoon animals, which integrates a graph neural network with a shape-aware relation learning module. To mitigate the significant differences among animals, a spatial and structural consistency learning strategy is proposed to capture and learn feature correlations across different animal species. Extensive experiments conducted on benchmark datasets demonstrate the effectiveness of the proposed approach, which outperforms the state-of-the-art methods.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get d3123f37-7dda-4004-89bc-54e9b5ff2edcRelated papers
- CPNet: Cartoon Parsing with Pixel and Part CorrelationJian-Jun Qiao, Jie Zhang, Xiao Wu, Yu-Pei Song et al.ACM MM 2023 · 2 citations
- CartoonNet: Cartoon Parsing with Semantic Consistency and Structure CorrelationJian-Jun Qiao, Meng-Yu Duan, Xiao Wu, Yu-Pei SongACM MM 2024
- Learning From Synthetic AnimalsJiteng Mu, Weichao Qiu, Gregory D. Hager, Alan L. YuilleCVPR 2020
- DAE-Net: Deforming Auto-Encoder for fine-grained shape co-segmentationZhiqin Chen, Qimin Chen, Hang Zhou, Hao ZhangSIGGRAPH 2024 · 9 citations
- Part-Aware Context Network for Human ParsingXiaomei Zhang, Yingying Chen, Bingke Zhu, Jinqiao Wang et al.CVPR 2020
