Learning Parallax Transformer Network for Stereo Image JPEG Artifacts Removal
Xuhao Jiang, Weimin Tan, Ri Cheng, Shili Zhou, Bo Yan
Abstract
Under stereo settings, the performance of image JPEG artifacts removal can be further improved by exploiting the additional information provided by a second view. However, incorporating this information for stereo image JPEG artifacts removal is a huge challenge, since the existing compression artifacts make pixel-level view alignment difficult. In this paper, we propose a novel parallax transformer network (PTNet) to integrate the information from stereo image pairs for stereo image JPEG artifacts removal. Specifically, a well-designed symmetric bi-directional parallax transformer module is proposed to match features with similar textures between different views instead of pixel-level view alignment. Due to the issues of occlusions and boundaries, a confidence-based cross-view fusion module is proposed to achieve better feature fusion for both views, where the cross-view features are weighted with confidence maps. Especially, we adopt a coarse-to-fine design for the cross-view interaction, leading to better performance. Comprehensive experimental results demonstrate that our PTNet can effectively remove compression artifacts and achieves superior performance than other testing state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Addressing Imbalance for Class Incremental Learning in Medical Image ClassificationXuze Hao, Wenqian Ni, Xuhao Jiang, Weimin Tan et al.ACM MM 2024 · 7 citations
- Uncover Treasures in DCT: Advancing JPEG Quality Enhancement by Exploiting Latent CorrelationsJing Yang, Qunliang Xing, Mai Xu, Minglang QiaoICCV 2025
Builds on9
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou et al.CVPR 2022 · 1,970 citations
- Towards Flexible Blind JPEG Artifacts RemovalJiaxi Jiang, Kai Zhang, Radu TimofteICCV 2021 · 145 citations
- JPEG Artifacts Reduction via Deep Convolutional Sparse CodingXueyang Fu, Zheng-Jun Zha, Feng Wu, Xinghao Ding et al.ICCV 2019 · 117 citations
Related papers
- Deep Stereo Video InpaintingZhiliang Wu, Changchang Sun, Hanyu Xuan, Yan YanCVPR 2023
- SIR-Former: Stereo Image Restoration Using TransformerZizheng Yang, Mingde Yao, Jie Huang, Man Zhou et al.ACM MM 2022 · 24 citations
- Learning Intra-View and Cross-View Geometric Knowledge for Stereo MatchingRui Gong, Weide Liu, Zaiwang Gu, Xulei Yang et al.CVPR 2024
- DLVINet: Advancing Dual-Lens Video Inpainting Beyond Parallax ConstraintsZhiliang Wu, Kun Li, Yunqiu Xu, Hehe Fan et al.AAAI 2026 · 1 citation
- Learning Pixel-wise Alignment for Unsupervised Image StitchingQi Jia, Xiaomei Feng, Yu Liu, Xin Fan et al.ACM MM 2023 · 34 citations
