Detail-Preserving Transformer for Light Field Image Super-resolution
Shunzhou Wang, Tianfei Zhou, Yao Lu, Huijun Di
Abstract
Recently, numerous algorithms have been developed to tackle the problem of light field super-resolution (LFSR), i.e., super-resolving low-resolution light fields to gain high-resolution views. Despite delivering encouraging results, these approaches are all convolution-based, and are naturally weak in global relation modeling of sub-aperture images necessarily to characterize the inherent structure of light fields. In this paper, we put forth a novel formulation built upon Transformers, by treating LFSR as a sequence-to-sequence reconstruction task. In particular, our model regards sub-aperture images of each vertical or horizontal angular view as a sequence, and establishes long-range geometric dependencies within each sequence via a spatial-angular locally-enhanced self-attention layer, which maintains the locality of each sub-aperture image as well. Additionally, to better recover image details, we propose a detail-preserving Transformer (termed as DPT), by leveraging gradient maps of light field to guide the sequence learning. DPT consists of two branches, with each associated with a Transformer for learning from an original or gradient image sequence. The two branches are finally fused to obtain comprehensive feature representations for reconstruction. Evaluations are conducted on a number of light field datasets, including real-world scenes and synthetic data. The proposed method achieves superior performance comparing with other state-of-the-art schemes. Our code is publicly available at: https://github.com/BITszwang/DPT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 03592bc2-dca0-47fa-91ab-5bff11a444b0Cited by top-tier papers4
- Learning Non-Local Spatial-Angular Correlation for Light Field Image Super-ResolutionZhengyu Liang, Yingqian Wang, Longguang Wang, Jungang Yang et al.ICCV 2023 · 72 citations
- GaTector: A Unified Framework for Gaze Object PredictionBinglu Wang, Tao Hu, Baoshan Li, Xiaojuan Chen et al.CVPR 2022 · 3 citations
- Rethinking the Upsampling Process in Light Field Super-Resolution with Spatial-Epipolar Implicit Image FunctionRuixuan Cong, Yu Wang, Mingyuan Zhao, Da Yang et al.ICCV 2025 · 1 citation
- CutMIB: Boosting Light Field Super-Resolution via Multi-View Image BlendingZeyu Xiao, Yutong Liu, Ruisheng Gao, Zhiwei XiongCVPR 2023
Builds on5
- Exploring Cross-Image Pixel Contrast for Semantic SegmentationWenguan Wang, Tianfei Zhou, Fisher Yu, Jifeng Dai et al.ICCV 2021 · 568 citations
- Learning Texture Transformer Network for Image Super-ResolutionFuzhi Yang, Huan Yang, Jianlong Fu, Hongtao Lu et al.CVPR 2020
- Pre-Trained Image Processing TransformerHanting Chen, Yunhe Wang, Tianyu Guo, Chang Xu et al.CVPR 2021
- Structure-Preserving Super Resolution With Gradient GuidanceCheng Ma, Yongming Rao, Yean Cheng, Ce Chen et al.CVPR 2020
- Light Field Spatial Super-Resolution via Deep Combinatorial Geometry Embedding and Structural Consistency RegularizationJing Jin, Junhui Hou, Jie Chen, Sam KwongCVPR 2020
Related papers
- Occlusion-Embedded Hybrid Transformer for Light Field Super-ResolutionZeyu Xiao, Zhuoyuan Li, Wei JiaAAAI 2025 · 24 citations
- Light Field Super-resolution via Attention-Guided Fusion of Hybrid LensesJing Jin, Junhui Hou, Jie Chen, Sam Kwong et al.ACM MM 2020 · 33 citations
- Learning Light Field Angular Super-Resolution via a Geometry-Aware NetworkJing Jin, Junhui Hou, Hui Yuan, Sam KwongAAAI 2020 · 124 citations
- From Coarse to Fine: Hierarchical Pixel Integration for Lightweight Image Super-resolutionJie Liu, Chao Chen, Jie Tang, Gangshan WuAAAI 2023 · 26 citations
- Light Field Super-Resolution With Zero-Shot LearningZhen Cheng, Zhiwei Xiong, Chang Chen, Dong Liu et al.CVPR 2021
