Learning Generalized Segmentation for Foggy-Scenes by Bi-directional Wavelet Guidance
Qi Bi, Shaodi You, Theo Gevers
Abstract
Learning scene semantics that can be well generalized to foggy conditions is important for safety-crucial applications such as autonomous driving. Existing methods need both annotated clear images and foggy images to train a curriculum domain adaptation model. Unfortunately, these methods can only generalize to the target foggy domain that has seen in the training stage, but the foggy domains vary a lot in both urban-scene styles and fog styles. In this paper, we propose to learn scene segmentation well generalized to foggy-scenes under the domain generalization setting, which does not involve any foggy images in the training stage and can generalize to any arbitrary unseen foggy scenes. We argue that an ideal segmentation model that can be well generalized to foggy-scenes need to simultaneously enhance the content, de-correlate the urban-scene style and de-correlate the fog style. As the content (e.g., scene semantic) rests more in low-frequency features while the style of urban-scene and fog rests more in high-frequency features, we propose a novel bi-directional wavelet guidance (BWG) mechanism to realize the above three objectives in a divide-and-conquer manner. With the aid of Haar wavelet transformation, the low frequency component is concentrated on the content enhancement self-attention, while the high frequency component is shifted to the style and fog self-attention for de-correlation purpose. It is integrated into existing mask-level Transformer segmentation pipelines in a learnable fashion. Large-scale experiments are conducted on four foggy-scene segmentation datasets under a variety of interesting settings. The proposed method significantly outperforms existing directly-supervised, curriculum domain adaptation and domain generalization segmentation methods. Source code is available at https://github.com/BiQiWHU/BWG.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 54cfde8c-c371-4ceb-8970-caffe269a4c1Cited by top-tier papers11
- Learning Frequency-Adapted Vision Foundation Model for Domain Generalized Semantic SegmentationQi Bi, Jingjun Yi, Hao Zheng, Haolan Zhan et al.NeurIPS 2024 · 62 citations
- Learning Spectral-Decomposited Tokens for Domain Generalized Semantic SegmentationJingjun Yi, Qi Bi, Hao Zheng, Haolan Zhan et al.ACM MM 2024 · 25 citations
- Samba: Severity-aware Recurrent Modeling for Cross-domain Medical Image GradingQi Bi, Jingjun Yi, Hao Zheng, Wei Ji et al.NeurIPS 2024 · 10 citations
- Learning Fine-grained Domain Generalization via Hyperbolic State Space HallucinationQi Bi, Jingjun Yi, Haolan Zhan, Wei Ji et al.AAAI 2025 · 8 citations
- Unleashing Multispectral Video's Potential in Semantic Segmentation: A Semi-supervised Viewpoint and New UAV-View BenchmarkWei Ji, Jingjing Li, Wenbo Li, Yilin Shen et al.NeurIPS 2024 · 8 citations
Builds on30
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- ACDC: The Adverse Conditions Dataset with Correspondences for Semantic Driving Scene UnderstandingChristos Sakaridis, Dengxin Dai, Luc Van GoolICCV 2021 · 655 citations
Related papers
- Both Style and Fog Matter: Cumulative Domain Adaptation for Semantic Foggy Scene UnderstandingXianzheng Ma, Zhixiang Wang, Yacheng Zhan, Yinqiang Zheng et al.CVPR 2022
- Learning Content-Enhanced Mask Transformer for Domain Generalized Urban-Scene SegmentationQi Bi, Shaodi You, Theo GeversAAAI 2024 · 77 citations
- FIFO: Learning Fog-invariant Features for Foggy Scene SegmentationSohyun Lee, Taeyoung Son, Suha KwakCVPR 2022 · 75 citations
- Adaptive Texture Filtering for Single-Domain Generalized SegmentationXinhui Li, Mingjia Li, Yaxing Wang, Chuan-Xian Ren et al.AAAI 2023 · 9 citations
- Train One, Generalize to All: Generalizable Semantic Segmentation from Single-Scene to All Adverse ScenesZiyang Gong, Fuhao Li, Yupeng Deng, Wenjun Shen et al.ACM MM 2023 · 9 citations
