Event-Customized Image Generation
Zhen Wang, Yilei Jiang, Dong Zheng, Jun Xiao, Long Chen
摘要
<V> sleeping <V> in a doghouse a girl wearing a hat and a scarf <V> Spiderman <V> a panda <V> <V> monkey <V> monkey cat <V> cat <V> girl , hat scarf <V1> <V2> ① <V1> ② <V2> ③ cat (a) The subject customization (b) The action and interaction customization (c) The event customization Figure 1: Customized Image Generation. (a) Generating customized images with given subjects in new contexts. (b) Generating customized images with co-existing basic action or interaction in given images. (c) Generating customized images for complex events with various target entities. Different colors and numbers show associations between reference entities and corresponding target prompts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- : Discrete Diffusion Model for Occluded 3D Human Pose EstimationWeiquan Wang, Jun Xiao, Chunping Wang, Wei Liu 等NeurIPS 2024 · 被引用 4 次
- SpA2V: Harnessing Spatial Auditory Cues for Audio-driven Spatially-aware Video GenerationKien T. Pham, Yingqing He, Yazhou Xing, Qifeng Chen 等ACM MM 2025 · 被引用 1 次
它引用的顶会 Paper32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- Learning Disentangled Identifiers for Action-Customized Text-to-Image GenerationSiteng Huang, Biao Gong, Yutong Feng, Xi Chen 等CVPR 2024
- I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion ModelsZhenxing Mi, Kuan-Chieh Wang, Guocheng Qian, Hanrong Ye 等ICML 2025
- Interact-Custom: Customized Human Object Interaction Image GenerationZhu Xu, Zhaowen Wang, Yuxin Peng, Yang LiuACM MM 2025 · 被引用 1 次
- TextCraftor: Your Text Encoder can be Image Quality ControllerYanyu Li, Xian Liu, Anil Kag, Ju Hu 等CVPR 2024
- Customize your NeRF: Adaptive Source Driven 3D Scene Editing via Local-Global Iterative TrainingRunze He, Shaofei Huang, Xuecheng Nie, Tianrui Hui 等CVPR 2024
