Diving Segmentation Model into Pixels
Chen Gan, Zihao Yin, Kelei He, Yang Gao, Junfeng Zhang
摘要
More distinguishable and consistent pixel features for each category will benefit the semantic segmentation under various settings. Existing efforts to mine better pixel-level features attempt to explicitly model the categorical distribution, which fails to achieve optimal due to the significant pixel feature variance. Moreover, prior research endeavors have scarcely delved into the thorough analysis and meticulous handling of pixel-level variance, leaving semantic segmentation at a coarse granularity. In this work, we analyze the causes of pixel-level variance and introduce the concept of pixel learning to concentrate on the tailored learning process of pixels to handle the pixel-level variance, enhancing the per-pixel recognition capability of segmentation models. Under the context of the pixel learning scheme, each image is viewed as a distribution of pixels, and pixel learning aims to pursue consistent pixel representation inside an image, continuously align pixels from different images (distributions), and eventually achieve consistent pixel representation for each category, even cross-domains. We proposed a pure pixel-level learning framework, namely PiXL, which consists of a pixel partition module to divide pixels into sub-domains, a prototype generation, a selection module to prepare targets for subsequent alignment, and a pixel alignment module to guarantee pixel feature consistency intra-/inter-images, and inter-domains. Extensive evaluations of multiple learning paradigms, including unsupervised domain adaptation and semi-/fully-supervised segmentation, show that PiXL outperforms state-ofthe-art performances, especially when annotated images are scarce. Visualization of the embedding space further demonstrates that pixel learning attains a superior representation of pixel features. The code is available here.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper24
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object DetectionHao Zhang, Feng Li, Shilong Liu, Lei Zhang 等ICLR 2023 · 被引用 753 次
- Exploring Cross-Image Pixel Contrast for Semantic SegmentationWenguan Wang, Tianfei Zhou, Fisher Yu, Jifeng Dai 等ICCV 2021 · 被引用 568 次
- DAFormer: Improving Network Architectures and Training Strategies for Domain-Adaptive Semantic SegmentationLukas Hoyer, Dengxin Dai, Luc Van GoolCVPR 2022 · 被引用 562 次
相关 Paper
- Pixel-level Intra-domain Adaptation for Semantic SegmentationZizheng Yan, Xianggang Yu, Yipeng Qin, Yushuang Wu 等ACM MM 2021 · 被引用 16 次
- SSA-Seg: Semantic and Spatial Adaptive Pixel-level Classifier for Semantic SegmentationXiaowen Ma, Zhenliang Ni, Xinghao ChenNeurIPS 2024
- PiPa: Pixel- and Patch-wise Self-supervised Learning for Domain Adaptative Semantic SegmentationMu Chen, Zhedong Zheng, Yi Yang, Tat-Seng ChuaACM MM 2023 · 被引用 65 次
- FedSeg: Class-Heterogeneous Federated Learning for Semantic SegmentationJiaxu Miao, Zongxin Yang, Leilei Fan, Yi YangCVPR 2023
- Pixel-Level Cycle Association: A New Perspective for Domain Adaptive Semantic SegmentationGuoliang Kang, Yunchao Wei, Yi Yang, Yueting Zhuang 等NeurIPS 2020 · 被引用 124 次
