SiamTrans: Zero-Shot Multi-Frame Image Restoration with Pre-trained Siamese Transformers
Lin Liu, Shanxin Yuan, Jianzhuang Liu, Xin Guo, Youliang Yan, Qi Tian
Abstract
We propose a novel zero-shot multi-frame image restoration method for removing unwanted obstruction elements (such as rains, snow, and moire patterns) that vary in successive frames. It has three stages: transformer pre-training, zero-shot restoration, and hard patch refinement. Using the pre-trained transformers, our model is able to tell the motion difference between the true image information and the obstructing elements. For zero-shot image restoration, we design a novel model, termed SiamTrans, which is constructed by Siamese transformers, encoders, and decoders. Each transformer has a temporal attention layer and several self-attention layers, to capture both temporal and spatial information of multiple frames. Only self-supervisedly pre-trained on the denoising task, SiamTrans is tested on three different low-level vision tasks (deraining, demoireing, and desnowing). Compared with related methods, SiamTrans achieves the best performances, even outperforming those with supervised learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bfe4e95d-49cd-4ea6-9479-498b1a88630fCited by top-tier papers2
- Improving Dynamic HDR Imaging with Fusion TransformerRufeng Chen, Bolun Zheng, Hua Zhang, Quan Chen et al.AAAI 2023 · 34 citations
- PAN-Crafter: Learning Modality-Consistent Alignment for Pan-SharpeningJeonghyeok Do, Sungpyo Kim, Geunhyuk Youk, Jaehyup Lee et al.ICCV 2025 · 3 citations
Builds on15
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Coherent Semantic Attention for Image InpaintingHongyu Liu, Bin Jiang, Yi Xiao, Chao YangICCV 2019 · 395 citations
- EfficientDeRain: Learning Pixel-wise Dilation Filtering for High-Efficiency Single-Image DerainingQing Guo, Jingyang Sun, Felix Juefei-Xu, Lei Ma et al.AAAI 2021 · 120 citations
- Joint Demosaicking and Denoising by Fine-Tuning of Bursts of Raw ImagesThibaud Ehret, Axel Davy, Pablo Arias, Gabriele FaccioloICCV 2019 · 54 citations
Related papers
- Close the Loop: A Unified Bottom-Up and Top-Down Paradigm for Joint Image Deraining and SegmentationYi Li, Yi Chang, Changfeng Yu, Luxin YanAAAI 2022 · 31 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- Siamese DETRZeren Chen, Gengshi Huang, Wei Li, Jianing Teng et al.CVPR 2023
- Zero-shot Video Restoration and Enhancement Using Pre-Trained Image Diffusion ModelCong Cao, Huanjing Yue, Xin Liu, Jingyu YangAAAI 2025 · 7 citations
- Video Adverse-Weather-Component Suppression Network via Weather Messenger and Adversarial BackpropagationYijun Yang, Angelica I. Avilés-Rivero, Huazhu Fu, Ye Liu et al.ICCV 2023 · 32 citations
