Wish You Were Here: Context-Aware Human Generation
Oran Gafni, Lior Wolf
Abstract
We present a novel method for inserting objects, specifically humans, into existing images, such that they blend in a photorealistic manner, while respecting the semantic context of the scene. Our method involves three subnetworks: the first generates the semantic map of the new person, given the pose of the other persons in the scene and an optional bounding box specification. The second network renders the pixels of the novel person and its blending mask, based on specifications in the form of multiple appearance components. A third network refines the generated face in order to match those of the target person. Our experiments present convincing high-resolution outputs in this novel and challenging application domain. In addition, the three networks are evaluated individually, demonstrating for example, state of the art results in pose transfer benchmarks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b1b1fdf5-003d-484a-8fa6-5cf974df4329Cited by top-tier papers11
- Image Manipulation Detection by Multi-View Multi-Scale SupervisionXinru Chen, Chengbo Dong, Jiaqi Ji, Juan Cao et al.ICCV 2021 · 271 citations
- TF-ICON: Diffusion-Based Training-Free Cross-Domain Image CompositionShilin Lu, Yanzhu Liu, Adams Wai-Kin KongICCV 2023 · 214 citations
- DiffEdit: Diffusion-based semantic image editing with mask guidanceGuillaume Couairon, Jakob Verbeek, Holger Schwenk, Matthieu CordICLR 2023 · 102 citations
- FlexIT: Towards Flexible Semantic Image TranslationGuillaume Couairon, Asya Grechka, Jakob Verbeek, Holger Schwenk et al.CVPR 2022 · 36 citations
- SDGAN: Disentangling Semantic Manipulation for Facial Attribute EditingWenmin Huang, Weiqi Luo, Jiwu Huang, Xiaochun CaoAAAI 2024 · 20 citations
Builds on2
Related papers
- Learning Semantic Person Image Generation by Region-Adaptive NormalizationZhengyao Lv, Xiaoming Li, Xin Li, Fu Li et al.CVPR 2021
- Single-Shot Freestyle Dance ReenactmentOran Gafni, Oron Ashual, Lior WolfCVPR 2021
- Human Synthesis and Scene CompositingMihai Zanfir, Elisabeta Oneata, Alin-Ionut Popa, Andrei Zanfir et al.AAAI 2020 · 24 citations
- Structure-aware Person Image Generation with Pose Decomposition and Semantic CorrelationJilin Tang, Yi Yuan, Tianjia Shao, Yong Liu et al.AAAI 2021 · 22 citations
- PISE: Person Image Synthesis and Editing With Decoupled GANJinsong Zhang, Kun Li, Yu-Kun Lai, Jingyu YangCVPR 2021
