CAPNet: Cartoon Animal Parsing with Spatial Learning and Structural Modeling
Jian-Jun Qiao, Meng-Yu Duan, Xiao Wu, Wei Li
摘要
Cartoon animal parsing aims to segment the body parts such as heads, arms, legs and tails of cartoon animals. Different from previous parsing tasks, cartoon animal parsing faces new challenges, including irregular body structures, abstract drawing styles and diverse animal categories. Existing methods have difficulties when addressing these challenges caused by the spatial and structural properties of cartoon animals. To address these challenges, a novel spatial learning and structural modeling network, named CAPNet, is proposed for cartoon animal parsing. It aims to address the critical problems of spatial perception, structure modeling and spatial-structural consistency learning. A spatial-aware learning module integrates deformable convolutions to learn spatial features of diverse cartoon animals. The multi-task edge and center point prediction mechanism is incorporated to capture the intricate spatial patterns. A structural modeling method is proposed to model the complex structural representations of cartoon animals, which integrates a graph neural network with a shape-aware relation learning module. To mitigate the significant differences among animals, a spatial and structural consistency learning strategy is proposed to capture and learn feature correlations across different animal species. Extensive experiments conducted on benchmark datasets demonstrate the effectiveness of the proposed approach, which outperforms the state-of-the-art methods.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- CPNet: Cartoon Parsing with Pixel and Part CorrelationJian-Jun Qiao, Jie Zhang, Xiao Wu, Yu-Pei Song 等ACM MM 2023 · 被引用 2 次
- CartoonNet: Cartoon Parsing with Semantic Consistency and Structure CorrelationJian-Jun Qiao, Meng-Yu Duan, Xiao Wu, Yu-Pei SongACM MM 2024
- Learning From Synthetic AnimalsJiteng Mu, Weichao Qiu, Gregory D. Hager, Alan L. YuilleCVPR 2020
- DAE-Net: Deforming Auto-Encoder for fine-grained shape co-segmentationZhiqin Chen, Qimin Chen, Hang Zhou, Hao ZhangSIGGRAPH 2024 · 被引用 9 次
- Part-Aware Context Network for Human ParsingXiaomei Zhang, Yingying Chen, Bingke Zhu, Jinqiao Wang 等CVPR 2020
