Exploring Dual-task Correlation for Pose Guided Person Image Generation
Pengze Zhang, Lingxiao Yang, Jianhuang Lai, Xiaohua Xie
Abstract
Pose Guided Person Image Generation (PGPIG) is the task of transforming a person image from the source pose to a given target pose. Most of the existing methods only focus on the ill-posed source-to-target task and fail to capture reasonable texture mapping. To address this problem, we propose a novel Dual-task Pose Transformer Network (DPTN), which introduces an auxiliary task (i.e., source-tosource task) and exploits the dual-task correlation to promote the performance of PGPIG. The DPTN is of a Siamese structure, containing a source-to-source self-reconstruction branch, and a transformation branch for source-to-target generation. By sharing partial weights between them, the knowledge learned by the source-to-source task can effectively assist the source-to-target learning. Furthermore, we bridge the two branches with a proposed Pose Transformer Module (PTM) to adaptively explore the correlation between features from dual tasks. Such correlation can establish the fine-grained mapping of all the pixels between the sources and the targets, and promote the source texture transmission to enhance the details of the generated target images. Extensive experiments show that our DPTN outperforms state-of-the-arts in terms of both PSNR and LPIPS. In addition, our DPTN only contains 9.79 million parameters, which is significantly smaller than other approaches. Our code is available at: https:// github.com/PangzeCheung/Dual-task-Pose- Transformer-Network.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e2c718c1-ffbe-48bd-93aa-2218d4e7e2dcCited by top-tier papers35
- IMAGPose: A Unified Conditional Framework for Pose-Guided Person GenerationFei Shen, Jinhui TangNeurIPS 2024 · 172 citations
- HumanSD: A Native Skeleton-Guided Diffusion Model for Human Image GenerationXuan Ju, Ailing Zeng, Chenchen Zhao, Jianan Wang et al.ICCV 2023 · 137 citations
- Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion ModelsFei Shen, Hu Ye, Jun Zhang, Cong Wang et al.ICLR 2024 · 133 citations
- HyperHuman: Hyper-Realistic Human Generation with Latent Structural DiffusionXian Liu, Jian Ren, Aliaksandr Siarohin, Ivan Skorokhodov et al.ICLR 2024 · 81 citations
- Controllable Person Image Synthesis with Pose-Constrained Latent DiffusionXiao Han, Xiatian Zhu, Jiankang Deng, Yi-Zhe Song et al.ICCV 2023 · 36 citations
Builds on7
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- Generative Adversarial TransformersDrew A. Hudson, Larry ZitnickICML 2021 · 213 citations
- PISE: Person Image Synthesis and Editing With Decoupled GANJinsong Zhang, Kun Li, Yu-Kun Lai, Jingyu YangCVPR 2021
Related papers
- Structure-transformed Texture-enhanced Network for Person Image SynthesisMunan Xu, Yuanqi Chen, Shan Liu, Thomas H. Li et al.ICCV 2021 · 3 citations
- Structure-aware Person Image Generation with Pose Decomposition and Semantic CorrelationJilin Tang, Yi Yuan, Tianjia Shao, Yong Liu et al.AAAI 2021 · 22 citations
- Learning Semantic Person Image Generation by Region-Adaptive NormalizationZhengyao Lv, Xiaoming Li, Xin Li, Fu Li et al.CVPR 2021
- Combining Attention with Flow for Person Image SynthesisYurui Ren, Yubo Wu, Thomas H. Li, Shan Liu et al.ACM MM 2021 · 16 citations
- Person Image Synthesis via Denoising Diffusion ModelAnkan Kumar Bhunia, Salman H. Khan, Hisham Cholakkal, Rao Muhammad Anwer et al.CVPR 2023
