Beyond Wide-Angle Images: Structure-to-Detail Video Portrait Correction via Unsupervised Spatiotemporal Adaptation
Wenbo Nie, Lang Nie, Chunyu Lin, Jingwen Chen, Ke Xing, Jiyuan Wang, Kang Liao
摘要
Wide-angle cameras, despite their popularity for content creation, suffer from distortion-induced facial stretching—especially at the edge of the lens—which degrades visual appeal. To address this issue, we propose a structure-to-detail portrait correction model named ImagePC. It integrates the long-range awareness of the transformer and multi-step denoising of diffusion models into a unified framework, achieving global structural robustness and local detail refinement. Besides, considering the high cost of obtaining video labels, we then repurpose ImagePC for unlabeled wide-angle videos (termed VideoPC), by spatiotemporal diffusion adaption with spatial consistency and temporal smoothness constraints. For the former, we encourage the denoised image to approximate pseudo labels following the wide-angle distortion distribution pattern, while for the latter, we derive rectification trajectories with backward optical flows and smooth them. Compared with ImagePC, VideoPC maintains high-quality facial corrections in space and mitigates the potential temporal shakes sequentially in blind scenarios. Finally, to establish an evaluation benchmark and train the framework, we establish a video portrait dataset with a large diversity in the number of people, lighting conditions, and background. Experiments demonstrate that the proposed methods outperform existing solutions quantitatively and qualitatively, contributing to high-fidelity wide-angle videos with stable and natural portraits.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan 等NeurIPS 2022 · 被引用 2,948 次
- Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video GenerationJay Zhangjie Wu, Yixiao Ge, Xintao Wang, Stan Weixian Lei 等ICCV 2023 · 被引用 1,113 次
- FateZero: Fusing Attentions for Zero-shot Text-based Video EditingChenyang Qi, Xiaodong Cun, Yong Zhang, Chenyang Lei 等ICCV 2023 · 被引用 510 次
- Sample and Computation Redistribution for Efficient Face DetectionJia Guo, Jiankang Deng, Alexandros Lattas, Stefanos ZafeiriouICLR 2022 · 被引用 173 次
- Fast Full-frame Video Stabilization with Iterative OptimizationWeiyue Zhao, Xin Li, Zhan Peng, Xianrui Luo 等ICCV 2023 · 被引用 24 次
相关 Paper
- Semi-Supervised Wide-Angle Portraits Correction by Multi-Scale TransformerFushun Zhu, Shan Zhao, Peng Wang, Hao Wang 等CVPR 2022 · 被引用 23 次
- Towards Complete Scene and Regular Shape for Distortion Rectification by Curve-Aware ExtrapolationKang Liao, Chunyu Lin, Yunchao Wei, Feng Li 等ICCV 2021 · 被引用 9 次
- Practical Wide-Angle Portraits Correction With Deep Structured ModelsJing Tan, Shan Zhao, Pengfei Xiong, Jiangyu Liu 等CVPR 2021
- RealPortrait: Realistic Portrait Animation with Diffusion TransformersZejun Yang, Huawei Wei, Zhisheng WangAAAI 2025 · 被引用 2 次
- Lifting the Structural Morphing for Wide-Angle Images Rectification: Unified Content and Boundary ModelingWenting Luan, Siqi Lu, Yongbin Zheng, Wanying Xu 等ICCV 2025 · 被引用 1 次
