Make-Your-Anchor: A Diffusion-based 2D Avatar Generation Framework
Ziyao Huang, Fan Tang, Yong Zhang, Xiaodong Cun, Juan Cao, Jintao Li, Tong-Yee Lee
Abstract
Video for model training. Video for model training. Motion condition Motion condition Generated anchor videos. Generated anchor videos. Figure 1 . We propose Make-Your-Anchor, a diffusion-based 2D avatar generation framework, to map the animatable human 3D mesh sequence into realistic human video. By combining with motion capturing or audio-to-motion methods, our framework could achieve anchor video auto-generation with temporal and lifelike outcomes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 526f9626-e1da-4b10-9053-42205088d62bCited by top-tier papers15
- ShowMaker: Creating High-Fidelity 2D Human Video via Fine-Grained Diffusion ModelingQuanwei Yang, Jiazhi Guan, Kaisiyuan Wang, Lingyun Yu et al.NeurIPS 2024 · 21 citations
- UltraGen: High-Resolution Video Generation with Hierarchical AttentionTeng Hu, Jiangning Zhang, Zihan Su, Ran YiAAAI 2026 · 7 citations
- Video Motion GraphsHaiyang Liu, Zhan Xu, Fa-Ting Hong, Hsin-Ping Huang et al.ICCV 2025 · 6 citations
- DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid GuidanceYuxuan Luo, Zhengkun Rong, Lizhen Wang, Longhao Zhang et al.ICCV 2025 · 5 citations
- A Multidimensional Measurement of Photorealistic Avatars Quality of ExperienceRoss Cutler, Babak Naderi, Vishak Gopal, Dharmendar Reddy PalleCSCW 2025 · 3 citations
Builds on20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- T2I-Adapter: Learning Adapters to Dig Out More Controllable Ability for Text-to-Image Diffusion ModelsChong Mou, Xintao Wang, Liangbin Xie, Yanze Wu et al.AAAI 2024 · 1,641 citations
Related papers
- Vid2Avatar-Pro: Authentic Avatar from Videos in the Wild via Universal PriorChen Guo, Junxuan Li, Yash Kant, Yaser Sheikh et al.CVPR 2025
- Move-in-2D: 2D-Conditioned Human Motion GenerationHsin-Ping Huang, Yang Zhou, Jui-Hsien Wang, Difan Liu et al.CVPR 2025
- GAS: Generative Avatar Synthesis from a Single ImageYixing Lu, Junting Dong, Youngjoong Kwon, Qin Zhao et al.ICCV 2025 · 5 citations
- Expressive Talking Human from Single-Image with Imperfect PriorsJun Xiang, Yudong Guo, Leipeng Hu, Boyang Guo et al.ICCV 2025 · 3 citations
- Make It Move: Controllable Image-to-Video Generation with Text DescriptionsYaosi Hu, Chong Luo, Zhenzhong ChenCVPR 2022 · 56 citations
