Occlusion-Embedded Hybrid Transformer for Light Field Super-Resolution
Zeyu Xiao, Zhuoyuan Li, Wei Jia
Abstract
Transformer-based networks have set new benchmarks in light field super-resolution (SR), but adapting them to capture both global and local spatial-angular correlations efficiently remains challenging. Moreover, many methods fail to account for geometric details like occlusions, leading to performance drops. To tackle these issues, we introduce OHT. This hybrid network leverages occlusion maps through an occlusionembedded mix layer. It combines the strengths of convolutional networks and Transformers via spatial-angular separable convolution (SASep-Conv) and angular self-attention (ASA). SASep-Conv offers a lightweight alternative to 3D convolution for capturing spatial-angular correlations, while the ASA mechanism applies 3D self-attention across the angular dimension. These designs allow OHT to capture global angular correlations effectively. Extensive experiments on multiple datasets demonstrate OHT's superior performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 996ce115-62b2-4994-88b4-183bf38d7d57Cited by top-tier papers5
- RewardMap: Tackling Sparse Rewards in Fine-grained Visual Reasoning via Multi-Stage Reinforcement LearningSicheng Feng, Kaiwen Tuo, Song Wang, Lingdong Kong et al.ICLR 2026 · 28 citations
- Hyperbolic Hierarchical Alignment Reasoning Network for Text-3D RetrievalWenrui Li, Yidan Lu, Yeyu Chai, Rui Zhao et al.AAAI 2026
- Exploiting Blurry Representations for Event-guided Video Super-ResolutionZeyu Xiao, Xinchao WangAAAI 2026
- Event-Guided Scene Text Image Super-ResolutionZihan Qi, Zeyu Xiao, Haoyi Zhao, Yang Zhao et al.AAAI 2026
- FreLay: Frequency-aware Energy Function for Training-free Layout-to-Image GenerationBonan Li, Yinhan Hu, Songhua Liu, Zeyu Xiao et al.AAAI 2026
Builds on11
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou et al.CVPR 2022 · 1,970 citations
- Learning Non-Local Spatial-Angular Correlation for Light Field Image Super-ResolutionZhengyu Liang, Yingqian Wang, Longguang Wang, Jungang Yang et al.ICCV 2023 · 72 citations
Related papers
- Detail-Preserving Transformer for Light Field Image Super-resolutionShunzhou Wang, Tianfei Zhou, Yao Lu, Huijun DiAAAI 2022 · 131 citations
- Spatial-angular Quality-aware Representation Learning for Blind Light Field Image Quality AssessmentJianjun Xiang, Yuanjie Dang, Peng Chen, Ronghua Liang et al.ACM MM 2023 · 4 citations
- Hybrid Spectral Denoising Transformer with Guided AttentionZeqiang Lai, Chenggang Yan, Ying FuICCV 2023 · 35 citations
- Learning Light Field Angular Super-Resolution via a Geometry-Aware NetworkJing Jin, Junhui Hou, Hui Yuan, Sam KwongAAAI 2020 · 124 citations
- Activating More Pixels in Image Super-Resolution TransformerXiangyu Chen, Xintao Wang, Jiantao Zhou, Yu Qiao et al.CVPR 2023
