OSRT: Omnidirectional Image Super-Resolution with Distortion-aware Transformer
Fanghua Yu, Xintao Wang, Mingdeng Cao, Gen Li, Ying Shan, Chao Dong
Abstract
Omnidirectional images (ODIs) have obtained lots of research interest for immersive experiences. Although ODIs require extremely high resolution to capture details of the entire scene, the resolutions of most ODIs are insufficient. Previous methods attempt to solve this issue by image super-resolution (SR) on equirectangular projection (ERP) images. However, they omit geometric properties of ERP in the degradation process, and their models can hardly generalize to real ERP images. In this paper, we propose Fisheye downsampling, which mimics the real-world imaging process and synthesizes more realistic low-resolution samples. Then we design a distortion-aware Transformer (OSRT) to modulate ERP distortions continuously and self-adaptively. Without a cumbersome process, OSRT outperforms previous methods by about 0.2dB on PSNR. Moreover, we propose a convenient data augmentation strategy, which synthesizes pseudo ERP images from plain images. This simple strategy can alleviate the over-fitting problem of large networks and significantly boost the performance of ODISR. Extensive experiments have demonstrated the state-of-theart performance of our OSRT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- CAMixerSR: Only Details Need More "Attention"Yan Wang, Yi Liu, Shijie Zhao, Junlin Li et al.CVPR 2024 · 60 citations
- Real-World Image Super-Resolution as Multi-Task LearningWenlong Zhang, Xiaohui Li, Guangyuan Shi, Xiangyu Chen et al.NeurIPS 2023 · 39 citations
- TULIP: Transformer for Upsampling of LiDAR Point CloudsBin Yang, Patrick Pfreundschuh, Roland Siegwart, Marco Hutter et al.CVPR 2024 · 18 citations
- Omnidirectional Image Super-resolution via Bi-projection FusionJiangang Wang, Yuning Cui, Yawen Li, Wenqi Ren et al.AAAI 2024 · 15 citations
- ResVR: Joint Rescaling and Viewport Rendering of Omnidirectional ImagesWeiqi Li, Shijie Zhao, Bin Chen, Xinhua Cheng et al.ACM MM 2024 · 6 citations
Builds on11
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Vision Transformer with Deformable AttentionZhuofan Xia, Xuran Pan, Shiji Song, Li Erran Li et al.CVPR 2022 · 835 citations
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- RankSRGAN: Generative Adversarial Networks With Ranker for Image Super-ResolutionWenlong Zhang, Yihao Liu, Chao Dong, Yu QiaoICCV 2019 · 406 citations
- Image Super-Resolution With Non-Local Sparse AttentionYiqun Mei, Yuchen Fan, Yuqian ZhouCVPR 2021
Related papers
- Fast Omni-Directional Image Super-Resolution: Adapting the Implicit Image Function with Pixel and Semantic-Wise Spherical Geometric PriorsXuelin Shen, Yitong Wang, Silin Zheng, Kang Xiao et al.AAAI 2025 · 3 citations
- Edge-Focused Super-Resolution for Omnidirectional Images with Spherical Geometric AugmentationShaolin Wang, Yuying Li, Lei Zhong, Shigang Li et al.CVPR 2026
- Spherical Pseudo-Cylindrical Representation for Omnidirectional Image Super-resolutionQing Cai, Mu Li, Dongwei Ren, Jun Lyu et al.AAAI 2024 · 11 citations
- S-OmniMVS: Incorporating Sphere Geometry into Omnidirectional Stereo MatchingZisong Chen, Chunyu Lin, Lang Nie, Zhijie Shen et al.ACM MM 2023 · 8 citations
- SimFIR: A Simple Framework for Fisheye Image Rectification with Self-supervised Representation LearningHao Feng, Wendi Wang, Jiajun Deng, Wengang Zhou et al.ICCV 2023 · 28 citations
