Tokenize Image Patches: Global Context Fusion for Effective Haze Removal in Large Images
Jiuchen Chen, Xinyu Yan, Qizhi Xu, Kaiqi Li
Abstract
Global contextual information and local detail features are essential for haze removal tasks. Deep learning models perform well on small, low-resolution images, but they encounter difficulties with large, high-resolution ones due to GPU memory limitations. As a compromise, they often resort to image slicing or downsampling. The former diminishes global information, while the latter discards highfrequency details. To address these challenges, we propose DehazeXL, a haze removal method that effectively balances global context and local feature extraction, enabling end-to-end modeling of large images on mainstream GPU hardware. Additionally, to evaluate the efficiency of global context utilization in haze removal performance, we design a visual attribution method tailored to the characteristics of haze removal tasks. Finally, recognizing the lack of benchmark datasets for haze removal in large images, we have developed an ultra-high-resolution haze removal dataset (8KDehaze) to support model training and testing. It includes 10000 pairs of clear and hazy remote sensing images, each sized at 8192 × 8192 pixels. Extensive experiments demonstrate that DehazeXL can infer images up to 10240 × 10240 pixels with only 21 GB of memory, achieving state-of-the-art results among all evaluated methods. The source code and experimental dataset are available at https://github.com/CastleChen339/ DehazeXL.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 80b067e1-2454-411f-9ccb-83fc215e8bd5Cited by top-tier papers2
- Fully Zero-Shot Image DehazingShuocheng Wang, Ruoxi Zhu, Jiaming Liu, Zhengyang Cao et al.ICML 2026
- Physically-Guided Optical Inversion Enable Non-Contact Side-Channel Attack on Isolated ScreensZhiwen Zheng, Yuheng Qiao, Xiaoshuai Zhang, Zhao Huang et al.ICLR 2026
Builds on15
- Swin Transformer V2: Scaling Up Capacity and ResolutionZe Liu, Han Hu, Yutong Lin, Zhuliang Yao et al.CVPR 2022 · 2,138 citations
- FFA-Net: Feature Fusion Attention Network for Single Image DehazingXu Qin, Zhilin Wang, Yuanchao Bai, Xiaodong Xie et al.AAAI 2020 · 1,828 citations
- Image Dehazing Transformer with Transmission-Aware 3D Position EmbeddingChunle Guo, Qixin Yan, Saeed Anwar, Runmin Cong et al.CVPR 2022 · 464 citations
- MB-TaylorFormer: Multi-branch Efficient Transformer Expanded by Taylor Formula for Image DehazingYuwei Qiu, Kaihao Zhang, Chenxi Wang, Wenhan Luo et al.ICCV 2023 · 224 citations
- HyperAttention: Long-context Attention in Near-Linear TimeInsu Han, Rajesh Jayaram, Amin Karbasi, Vahab Mirrokni et al.ICLR 2024 · 104 citations
Related papers
- Ultra-High-Definition Image Dehazing via Multi-Guided Bilateral LearningZhuoran Zheng, Wenqi Ren, Xiaochun Cao, Xiaobin Hu et al.CVPR 2021
- HazeSpace2M: A Dataset for Haze Aware Single Image DehazingMd Tanvir Islam, Nasir Rahim, Saeed Anwar, Muhammad Saqib et al.ACM MM 2024 · 19 citations
- DehazeFlow: Multi-scale Conditional Flow Network for Single Image DehazingHongyu Li, Jia Li, Dong Zhao, Long XuACM MM 2021 · 49 citations
- Video Dehazing via a Multi-Range Temporal Alignment Network with Physical PriorJiaqi Xu, Xiaowei Hu, Lei Zhu, Qi Dou et al.CVPR 2023
- Discrete Haze Level Dehazing NetworkXiaofeng Cong, Jie Gui, Kai-Chao Miao, Jun Zhang et al.ACM MM 2020 · 29 citations
