MagicCartoon: 3D Pose and Shape Estimation for Bipedal Cartoon Characters
Yu-Pei Song, Yuantong Liu, Xiao Wu, Qi He, Zhaoquan Yuan, Ao Luo
Abstract
The 3D model can be estimated by regressing the pose and shape parameters from the image data of the digital model. The reconstruction of 3D cartoon characters poses a challenging task due to diverse visual representations and postural variations. This paper proposes a dual-branch structure named MagicCartoon for 3D bipedal cartoon character estimation, which models pose and shape independently through feature decoupling. Considering the correlation between category difference and shape parameters, a hybrid feature fusion technique is introduced, which integrates the global features of the original image with the corresponding local features expressed by the puzzle image, reducing the abstractness of understanding shape parameter differences. To semantically align image and geometric between feature space, a geometric-guided feedback loop is proposed in an iterative way, so that the pose of modeling results can be expressed consistently with the image. Moreover, a feature consistency loss is designed to augment the training data by incorporating the same character with different postures and the same posture of different characters. It enhances the correlation between the features extracted by the backbone network and the specific task. Experiments conducted on the 3DBiCar dataset demonstrate that MagicCartoon outperforms the state-of-the-art methods.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 78813bce-4eca-4479-b2b6-afe7de88c637Cited by top-tier papers2
- Dehallu3D: Hallucination-Mitigated 3D Generation from a Single Image via Cyclic View Consistency RefinementXiwen Wang, Shichao Zhang, Ruowei Wang, Mao Li et al.CVPR 2026
- REVIVE 3D: Refinement via Encoded Voluminous Inflated prior for Volume EnhancementHankyeol Lee, Wooyeol Baek, Seongdo Kim, Jongyoo KimCVPR 2026
Related papers
- RaBit: Parametric Modeling of 3D Biped Cartoon Characters with a Topological-Consistent DatasetZhongjin Luo, Shengcai Cai, Jinguo Dong, Ruibo Ming et al.CVPR 2023
- Cartoon Face Recognition: A Benchmark DatasetYi Zheng, Yifan Zhao, Mengyuan Ren, He Yan et al.ACM MM 2020 · 53 citations
- HybrIK: A Hybrid Analytical-Neural Inverse Kinematics Solution for 3D Human Pose and Shape EstimationJiefeng Li, Chao Xu, Zhicun Chen, Siyuan Bian et al.CVPR 2021
- CharacterGen: Efficient 3D Character Generation from Single Images with Multi-View Pose CanonicalizationHao-Yang Peng, Jia-Peng Zhang, Meng-Hao Guo, Yan-Pei Cao et al.SIGGRAPH 2024 · 30 citations
- CartoonNet: Cartoon Parsing with Semantic Consistency and Structure CorrelationJian-Jun Qiao, Meng-Yu Duan, Xiao Wu, Yu-Pei SongACM MM 2024
