PISE: Person Image Synthesis and Editing With Decoupled GAN
Jinsong Zhang, Kun Li, Yu-Kun Lai, Jingyu Yang
Abstract
Person image synthesis, e.g., pose transfer, is a challenging problem due to large variation and occlusion. Existing methods have difficulties predicting reasonable invisible regions and fail to decouple the shape and style of clothing, which limits their applications on person image editing. In this paper, we propose PISE, a novel two-stage generative model for Person Image Synthesis and Editing, which is able to generate realistic person images with desired poses, textures, or semantic layouts. For human pose transfer, we first synthesize a human parsing map aligned with the target pose to represent the shape of clothing by a parsing generator, and then generate the final image by an image generator. To decouple the shape and style of clothing, we propose joint global and local per-region encoding and normalization to predict the reasonable style of clothing for invisible regions. We also propose spatial-aware normalization to retain the spatial context relationship in the source image. The results of qualitative and quantitative experiments demonstrate the superiority of our model on human pose transfer. Besides, the results of texture transfer and region editing show that our model can be applied to person image editing. The code is available for research purposes at https://github.com/Zhangjinso/PISE .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 261b4cc2-3dd2-4d34-a147-aded36462e84Cited by top-tier papers30
- DreamPose: Fashion Image-to-Video Synthesis via Stable DiffusionJohanna Suvi Karras, Aleksander Holynski, Ting-Chun Wang, Ira Kemelmacher-ShlizermanICCV 2023 · 224 citations
- IMAGPose: A Unified Conditional Framework for Pose-Guided Person GenerationFei Shen, Jinhui TangNeurIPS 2024 · 172 citations
- HumanSD: A Native Skeleton-Guided Diffusion Model for Human Image GenerationXuan Ju, Ailing Zeng, Chenchen Zhao, Jianan Wang et al.ICCV 2023 · 137 citations
- Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion ModelsFei Shen, Hu Ye, Jun Zhang, Cong Wang et al.ICLR 2024 · 133 citations
- Exploring Dual-task Correlation for Pose Guided Person Image GenerationPengze Zhang, Lingxiao Yang, Jianhuang Lai, Xiaohua XieCVPR 2022 · 92 citations
Builds on5
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen et al.ICCV 2019 · 1,990 citations
- StructureFlow: Image Inpainting via Structure-Aware Appearance FlowYurui Ren, Xiaoming Yu, Ruonan Zhang, Thomas H. Li et al.ICCV 2019 · 356 citations
- SEAN: Image Synthesis With Semantic Region-Adaptive NormalizationPeihao Zhu, Rameen Abdal, Yipeng Qin, Peter WonkaCVPR 2020
- Semantically Multi-Modal Image SynthesisZhen Zhu, Zhiliang Xu, Ansheng You, Xiang BaiCVPR 2020
- Controllable Person Image Synthesis With Attribute-Decomposed GANYifang Men, Yiming Mao, Yuning Jiang, Wei-Ying Ma et al.CVPR 2020
Related papers
- Learning Semantic Person Image Generation by Region-Adaptive NormalizationZhengyao Lv, Xiaoming Li, Xin Li, Fu Li et al.CVPR 2021
- Structure-transformed Texture-enhanced Network for Person Image SynthesisMunan Xu, Yuanqi Chen, Shan Liu, Thomas H. Li et al.ICCV 2021 · 3 citations
- Combining Attention with Flow for Person Image SynthesisYurui Ren, Yubo Wu, Thomas H. Li, Shan Liu et al.ACM MM 2021 · 16 citations
- PICTURE: PhotorealistIC Virtual Try-on from UnconstRained dEsignsShuliang Ning, Duomin Wang, Yipeng Qin, Zirong Jin et al.CVPR 2024 · 10 citations
- Human Parsing Based Texture Transfer from Single Image to 3D Human via Cross-View ConsistencyFang Zhao, Shengcai Liao, Kaihao Zhang, Ling ShaoNeurIPS 2020 · 22 citations
