Decoupled Finetuning for Domain Generalizable Semantic Segmentation
Jaehyun Pahk, Donghyeon Kwon, Seong Joon Oh, Suha Kwak
Abstract
Joint finetuning of a pretrained encoder and a randomly initialized decoder has been the de facto standard in semantic segmentation, but the vulnerability of this approach to domain shift has not been studied. We investigate the vulnerability issue of joint finetuning, and propose a novel finetuning framework called Decoupled FineTuning (DeFT) for domain generalization as a solution. DeFT operates in two stages. Its first stage warms up the decoder with the frozen, pretrained encoder so that the decoder learns task-relevant knowledge while the encoder preserves its generalizable features. In the second stage, it decouples finetuning of the encoder and decoder into two pathways, each of which concatenates an adaptive component (AC) and retentive component (RC); the encoder and decoder play different roles between AC and RC in different pathways. ACs are updated by gradients of the loss on the source domain, while RCs are updated by exponential moving average biased toward their initialization to retain their generalization capability. By the two separate optimization pathways with opposite AC-RC configurations, DeFT reduces the number of learnable parameters virtually, and decreases the distance between learned parameters and their initialization, leading to improved generalization capability. DeFT significantly outperformed existing methods in various domain shift scenarios, and its performance was further boosted by incorporating a simple distance regularization.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fad66d7b-e6c4-4f9f-9d4e-7cd5e94de8ddBuilds on25
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs et al.ICML 2022 · 1,464 citations
- MetaFormer is Actually What You Need for VisionWeihao Yu, Mi Luo, Pan Zhou, Chenyang Si et al.CVPR 2022 · 1,114 citations
- Domain Generalization with MixStyleKaiyang Zhou, Yongxin Yang, Yu Qiao, Tao XiangICLR 2021 · 986 citations
- Fine-Tuning can Distort Pretrained Features and Underperform Out-of-DistributionAnanya Kumar, Aditi Raghunathan, Robbie Matthew Jones, Tengyu Ma et al.ICLR 2022 · 911 citations
Related papers
- FisherTune: Fisher-Guided Robust Tuning of Vision Foundation Models for Domain Generalized SegmentationDong Zhao, Jinlong Li, Shuang Wang, Mengyao Wu et al.CVPR 2025
- Decoupling Continual Semantic SegmentationYifu Guo, Yuquan Lu, Wentao Zhang, Zishan Xu et al.AAAI 2026 · 3 citations
- Adapter Naturally Serves as Decoupler for Cross-Domain Few-Shot Semantic SegmentationJintao Tong, Ran Ma, Yixiong Zou, Guangyao Chen et al.ICML 2025
- GeCo: Geometry-Consistent Regularization for Domain Generalized Semantic SegmentationQi Zang, Dong Zhao, Nan Pu, Wenjing Li et al.CVPR 2026
- Stronger, Steadier & Superior: Geometric Consistency in Depth VFM Forges Domain Generalized Semantic SegmentationSiyu Chen, Ting Han, Changshe Zhang, Xin Luo et al.ICCV 2025 · 4 citations
