Learning Non-Local Spatial-Angular Correlation for Light Field Image Super-Resolution
Zhengyu Liang, Yingqian Wang, Longguang Wang, Jungang Yang, Shilin Zhou, Yulan Guo
Abstract
Exploiting spatial-angular correlation is crucial to light field (LF) image super-resolution (SR), but is highly challenging due to its non-local property caused by the disparities among LF images. Although many deep neural networks (DNNs) have been developed for LF image SR and achieved continuously improved performance, existing methods cannot well leverage the long-range spatial-angular correlation and thus suffer a significant performance drop when handling scenes with large disparity variations. In this paper, we propose a simple yet effective method to learn the non-local spatial-angular correlation for LF image SR. In our method, we adopt the epipolar plane image (EPI) representation to project the 4D spatial-angular correlation onto multiple 2D EPI planes, and then develop a Transformer network with repetitive self-attention operations to learn the spatial-angular correlation by modeling the dependencies between each pair of EPI pixels. Our method can fully incorporate the information from all angular views while achieving a global receptive field along the epipolar line. We conduct extensive experiments with insightful visualizations to validate the effectiveness of our method. Comparative results on five public datasets show that our method not only achieves state-of-the-art SR performance but also performs robust to disparity variations. Code is publicly available at https://github.com/ZhengyuLiang24/EPIT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1bd161b5-6859-413a-9e07-0dc6a70144f9Cited by top-tier papers3
- Occlusion-Embedded Hybrid Transformer for Light Field Super-ResolutionZeyu Xiao, Zhuoyuan Li, Wei JiaAAAI 2025 · 24 citations
- Rethinking the Upsampling Process in Light Field Super-Resolution with Spatial-Epipolar Implicit Image FunctionRuixuan Cong, Yu Wang, Mingyuan Zhao, Da Yang et al.ICCV 2025 · 1 citation
- CutMIB: Boosting Light Field Super-Resolution via Multi-View Image BlendingZeyu Xiao, Yutong Liu, Ruisheng Gao, Zhiwei XiongCVPR 2023
Builds on19
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- Video Swin TransformerZe Liu, Jia Ning, Yue Cao, Yixuan Wei et al.CVPR 2022 · 1,847 citations
- On Layer Normalization in the Transformer ArchitectureRuibin Xiong, Yunchang Yang, Di He, Kai Zheng et al.ICML 2020 · 1,388 citations
- Intriguing Properties of Vision TransformersMuzammal Naseer, Kanchana Ranasinghe, Salman Khan, Munawar Hayat et al.NeurIPS 2021 · 863 citations
Related papers
- Detail-Preserving Transformer for Light Field Image Super-resolutionShunzhou Wang, Tianfei Zhou, Yao Lu, Huijun DiAAAI 2022 · 131 citations
- Spatial-angular Quality-aware Representation Learning for Blind Light Field Image Quality AssessmentJianjun Xiang, Yuanjie Dang, Peng Chen, Ronghua Liang et al.ACM MM 2023 · 4 citations
- Epipolar Consistent Attention Aggregation Network for Unsupervised Light Field Disparity EstimationChen Gao, Shuo Zhang, Youfang LinICCV 2025 · 3 citations
- Epipolar Consistency-based Network for Structure-Aware LF Semantic SegmentationChen Gao, Youfang Lin, Wenbin Wang, Shuo ZhangACM MM 2025
- Combining Implicit-Explicit View Correlation for Light Field Semantic SegmentationRuixuan Cong, Da Yang, Rongshan Chen, Sizhe Wang et al.CVPR 2023
