Learning Disentangled Identifiers for Action-Customized Text-to-Image Generation
Siteng Huang, Biao Gong, Yutong Feng, Xi Chen, Yuqian Fu, Yu Liu, Donglin Wang
Abstract
<A> "A boy <A>" "Spiderman <A>" "Messi <A>" "A gorilla <A>" "A bear <A>" "A panda <A>" Sample Images "An old man <A>" "Batman <A>" <A> "Barack Obama <A>" "A monkey <A>" "A polar bear <A>" "A cat <A>" Sample Images <A> "A woman <A>" "David Beckham <A>" "Michael Jackson <A>" "A dog <A>" "A fox <A>" "A cheetah <A>" Sample Images Figure 1 . Action customization results of our ADI method. By inverting representative action-related features, the learned identifiers "<A>" can be paired with a variety of characters and animals to contribute to the generation of accurate, diverse and high-quality images.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 01a5918d-151d-4029-afa0-b3fad2ced523Cited by top-tier papers15
- Identity Decoupling for Multi-Subject Personalization of Text-to-Image ModelsSangwon Jang, Jaehyeong Jo, Kimin Lee, Sung Ju HwangNeurIPS 2024 · 42 citations
- ResMaster: Mastering High-Resolution Image Generation via Structural and Fine-Grained GuidanceShuwei Shi, Wenbo Li, Yuechen Zhang, Jingwen He et al.AAAI 2025 · 23 citations
- Action Imitation in Common Action Space for Customized Action Image SynthesisWang Lin, Jingyuan Chen, Jiaxin Shi, Zirun Guo et al.NeurIPS 2024 · 17 citations
- Lay2Story: Extending Diffusion Transformers for Layout-Togglable Story GenerationAo Ma, Jiasong Feng, Ke Cao, Jing Wang et al.ICCV 2025 · 13 citations
- Seg2Any: Open-set Segmentation-Mask-to-Image Generation with Precise Shape and Semantic ControlDanfeng Li, Hui Zhang, Sheng Wang, Jiacheng Li et al.NeurIPS 2025 · 11 citations
Builds on18
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
Related papers
- Event-Customized Image GenerationZhen Wang, Yilei Jiang, Dong Zheng, Jun Xiao et al.ICML 2025
- Articulated Kinematics Distillation from Video Diffusion ModelsXuan Li, Qianli Ma, Tsung-Yi Lin, Yongxin Chen et al.CVPR 2025
- CapHuman: Capture Your Moments in Parallel UniversesChao Liang, Fan Ma, Linchao Zhu, Yingying Deng et al.CVPR 2024
- IMAGINE: Image Synthesis by Image-Guided Model InversionPei Wang, Yijun Li, Krishna Kumar Singh, Jingwan Lu et al.CVPR 2021
- FlexiAct: Towards Flexible Action Control in Heterogeneous ScenariosShiyi Zhang, Junhao Zhuang, Zhaoyang Zhang, Ying Shan et al.SIGGRAPH 2025 · 8 citations
