OSRT: Omnidirectional Image Super-Resolution with Distortion-aware Transformer
Fanghua Yu, Xintao Wang, Mingdeng Cao, Gen Li, Ying Shan, Chao Dong
摘要
Omnidirectional images (ODIs) have obtained lots of research interest for immersive experiences. Although ODIs require extremely high resolution to capture details of the entire scene, the resolutions of most ODIs are insufficient. Previous methods attempt to solve this issue by image super-resolution (SR) on equirectangular projection (ERP) images. However, they omit geometric properties of ERP in the degradation process, and their models can hardly generalize to real ERP images. In this paper, we propose Fisheye downsampling, which mimics the real-world imaging process and synthesizes more realistic low-resolution samples. Then we design a distortion-aware Transformer (OSRT) to modulate ERP distortions continuously and self-adaptively. Without a cumbersome process, OSRT outperforms previous methods by about 0.2dB on PSNR. Moreover, we propose a convenient data augmentation strategy, which synthesizes pseudo ERP images from plain images. This simple strategy can alleviate the over-fitting problem of large networks and significantly boost the performance of ODISR. Extensive experiments have demonstrated the state-of-theart performance of our OSRT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- CAMixerSR: Only Details Need More "Attention"Yan Wang, Yi Liu, Shijie Zhao, Junlin Li 等CVPR 2024 · 被引用 60 次
- Real-World Image Super-Resolution as Multi-Task LearningWenlong Zhang, Xiaohui Li, Guangyuan Shi, Xiangyu Chen 等NeurIPS 2023 · 被引用 39 次
- TULIP: Transformer for Upsampling of LiDAR Point CloudsBin Yang, Patrick Pfreundschuh, Roland Siegwart, Marco Hutter 等CVPR 2024 · 被引用 18 次
- Omnidirectional Image Super-resolution via Bi-projection FusionJiangang Wang, Yuning Cui, Yawen Li, Wenqi Ren 等AAAI 2024 · 被引用 15 次
- ResVR: Joint Rescaling and Viewport Rendering of Omnidirectional ImagesWeiqi Li, Shijie Zhao, Bin Chen, Xinhua Cheng 等ACM MM 2024 · 被引用 6 次
它引用的顶会 Paper11
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Vision Transformer with Deformable AttentionZhuofan Xia, Xuran Pan, Shiji Song, Li Erran Li 等CVPR 2022 · 被引用 835 次
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 被引用 522 次
- RankSRGAN: Generative Adversarial Networks With Ranker for Image Super-ResolutionWenlong Zhang, Yihao Liu, Chao Dong, Yu QiaoICCV 2019 · 被引用 406 次
- Image Super-Resolution With Non-Local Sparse AttentionYiqun Mei, Yuchen Fan, Yuqian ZhouCVPR 2021
相关 Paper
- Fast Omni-Directional Image Super-Resolution: Adapting the Implicit Image Function with Pixel and Semantic-Wise Spherical Geometric PriorsXuelin Shen, Yitong Wang, Silin Zheng, Kang Xiao 等AAAI 2025 · 被引用 3 次
- Edge-Focused Super-Resolution for Omnidirectional Images with Spherical Geometric AugmentationShaolin Wang, Yuying Li, Lei Zhong, Shigang Li 等CVPR 2026
- Spherical Pseudo-Cylindrical Representation for Omnidirectional Image Super-resolutionQing Cai, Mu Li, Dongwei Ren, Jun Lyu 等AAAI 2024 · 被引用 11 次
- S-OmniMVS: Incorporating Sphere Geometry into Omnidirectional Stereo MatchingZisong Chen, Chunyu Lin, Lang Nie, Zhijie Shen 等ACM MM 2023 · 被引用 8 次
- SimFIR: A Simple Framework for Fisheye Image Rectification with Self-supervised Representation LearningHao Feng, Wendi Wang, Jiajun Deng, Wengang Zhou 等ICCV 2023 · 被引用 28 次
