Event-Customized Image Generation
Zhen Wang, Yilei Jiang, Dong Zheng, Jun Xiao, Long Chen
Abstract
<V> sleeping <V> in a doghouse a girl wearing a hat and a scarf <V> Spiderman <V> a panda <V> <V> monkey <V> monkey cat <V> cat <V> girl , hat scarf <V1> <V2> ① <V1> ② <V2> ③ cat (a) The subject customization (b) The action and interaction customization (c) The event customization Figure 1: Customized Image Generation. (a) Generating customized images with given subjects in new contexts. (b) Generating customized images with co-existing basic action or interaction in given images. (c) Generating customized images for complex events with various target entities. Different colors and numbers show associations between reference entities and corresponding target prompts.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext badfee02-7c7b-4e02-a54e-4694d2b5c565Cited by top-tier papers2
- : Discrete Diffusion Model for Occluded 3D Human Pose EstimationWeiquan Wang, Jun Xiao, Chunping Wang, Wei Liu et al.NeurIPS 2024 · 4 citations
- SpA2V: Harnessing Spatial Auditory Cues for Audio-driven Spatially-aware Video GenerationKien T. Pham, Yingqing He, Yazhou Xing, Qifeng Chen et al.ACM MM 2025 · 1 citation
Builds on32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- Learning Disentangled Identifiers for Action-Customized Text-to-Image GenerationSiteng Huang, Biao Gong, Yutong Feng, Xi Chen et al.CVPR 2024
- I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion ModelsZhenxing Mi, Kuan-Chieh Wang, Guocheng Qian, Hanrong Ye et al.ICML 2025
- Interact-Custom: Customized Human Object Interaction Image GenerationZhu Xu, Zhaowen Wang, Yuxin Peng, Yang LiuACM MM 2025 · 1 citation
- TextCraftor: Your Text Encoder can be Image Quality ControllerYanyu Li, Xian Liu, Anil Kag, Ju Hu et al.CVPR 2024
- Customize your NeRF: Adaptive Source Driven 3D Scene Editing via Local-Global Iterative TrainingRunze He, Shaofei Huang, Xuecheng Nie, Tianrui Hui et al.CVPR 2024
