Semi-Supervised Wide-Angle Portraits Correction by Multi-Scale Transformer
Fushun Zhu, Shan Zhao, Peng Wang, Hao Wang, Hua Yan, Shuaicheng Liu
摘要
We propose a semi-supervised network for wide-angle portraits correction. Wide-angle images often suffer from skew and distortion affected by perspective distortion, especially noticeable at the face regions. Previous deep learning based approaches need the ground-truth correction flow maps for training guidance. However, such labels are expensive, which can only be obtained manually. In this work, we design a semi-supervised scheme and build a high-quality unlabeled dataset with rich scenarios, allowing us to simultaneously use labeled and unlabeled data to improve performance. Specifically, our semi-supervised scheme takes advantage of the consistency mechanism, with several novel components such as direction and range consistency (DRC) and regression consistency (RC). Furthermore, different from the existing methods, we propose the Multi-Scale Swin-Unet (MS-Unet) based on the multi-scale swin transformer block (MSTB), which can simultaneously learn short-distance and long-distance information to avoid artifacts. Extensive experiments demonstrate that the proposed method is superior to the state-of-the-art methods and other representative baselines. The source code and dataset are available at https://github.com/megvii- research/Portraits_Correction
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- RecDiffusion: Rectangling for Image Stitching with Diffusion ModelsTianhao Zhou, Haipeng Li, Ziyi Wang, Ao Luo 等CVPR 2024 · 被引用 21 次
- Wide-Angle Rectification via Content-Aware Conformal MappingQi Zhang, Hongdong Li, Qing WangCVPR 2023
- MaDCoW: Marginal Distortion Correction for Wide-Angle Photography with Arbitrary ObjectsKevin Zhang, Jia-Bin Huang, Jose Echevarria, Stephen DiVerdi 等CVPR 2025
- Rectification Reimagined: A Unified Mamba Model for Image Correction and Rectangling with PromptsLinwei Qiu, Gongzhe Li, Xiaozhe Zhang, Qi Sun 等AAAI 2026
- Distilling Quasi-Conformal Mapping: A Generalizable and Efficient Solution for Wide-Angle CorrectionChengyang Liu, Zixuan Lin, Miaolin Han, Michael K. Ng 等CVPR 2026
它引用的顶会 Paper7
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- ResT: An Efficient Transformer for Visual RecognitionQinglong Zhang, Yu-Bin YangNeurIPS 2021 · 被引用 313 次
- Spatial Uncertainty-Aware Semi-Supervised Crowd CountingYanda Meng, Hongrun Zhang, Yitian Zhao, Xiaoyun Yang 等ICCV 2021 · 被引用 108 次
- SALNet: Semi-supervised Few-Shot Text Classification with Attention-based Lexicon ConstructionJu Hyoung Lee, Sang-Ki Ko, Yo-Sub HanAAAI 2021 · 被引用 26 次
相关 Paper
- Practical Wide-Angle Portraits Correction With Deep Structured ModelsJing Tan, Shan Zhao, Pengfei Xiong, Jiangyu Liu 等CVPR 2021
- Beyond Wide-Angle Images: Structure-to-Detail Video Portrait Correction via Unsupervised Spatiotemporal AdaptationWenbo Nie, Lang Nie, Chunyu Lin, Jingwen Chen 等AAAI 2026
- Dense Keypoints via Multiview SupervisionZhixuan Yu, Haozheng Yu, Long Sha, Sujoy Ganguly 等NeurIPS 2021
- DarSwin: Distortion Aware Radial Swin TransformerAkshaya Athwale, Arman Afrasiyabi, Justin Lagüe, Ichrak Shili 等ICCV 2023 · 被引用 13 次
- Learning Perspective Undistortion of PortraitsYajie Zhao, Zeng Huang, Tianye Li, Weikai Chen 等ICCV 2019 · 被引用 29 次
