GMMSeg: Gaussian Mixture based Generative Semantic Segmentation Models
Chen Liang, Wenguan Wang, Jiaxu Miao, Yi Yang
摘要
Prevalent semantic segmentation solutions are, in essence, a dense discriminative classifier of p(class |pixel feature). Though straightforward, this de facto paradigm neglects the underlying data distribution p(pixel feature |class), and struggles to identify out-of-distribution data. Going beyond this, we propose GMMSeg, a new family of segmentation models that rely on a dense generative classifier for the joint distribution p(pixel feature, class). For each class, GMMSeg builds Gaussian Mixture Models (GMMs) via Expectation-Maximization (EM), so as to capture class-conditional densities. Meanwhile, the deep dense representation is end-to-end trained in a discriminative manner, i.e., maximizing p(class |pixel feature). This endows GMMSeg with the strengths of both generative and discriminative models. With a variety of segmentation architectures and backbones, GMMSeg outperforms the discriminative counterparts on three closed-set datasets. More impressively, without any modification, GMMSeg even performs well on open-world datasets. We believe this work brings fundamental insights into the related fields.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper44
- MomentDiff: Generative Video Moment Retrieval from Random to RealPandeng Li, Chen-Wei Xie, Hongtao Xie, Liming Zhao 等NeurIPS 2023 · 被引用 113 次
- DiffusionRet: Generative Text-Video Retrieval with Diffusion ModelPeng Jin, Hao Li, Zesen Cheng, Kehan Li 等ICCV 2023 · 被引用 95 次
- CLUSTSEG: Clustering for Universal SegmentationJames Chenhao Liang, Tianfei Zhou, Dongfang Liu, Wenguan WangICML 2023 · 被引用 85 次
- RbA: Segmenting Unknown Regions Rejected by AllNazir Nayal, Misra Yavuz, João F. Henriques, Fatma GüneyICCV 2023 · 被引用 73 次
- Residual Pattern Learning for Pixel-wise Out-of-Distribution Detection in Semantic SegmentationYuyuan Liu, Choubo Ding, Yu Tian, Guansong Pang 等ICCV 2023 · 被引用 69 次
它引用的顶会 Paper38
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang 等ICCV 2019 · 被引用 2,972 次
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 被引用 2,213 次
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
相关 Paper
- Generative Semantic SegmentationJiaqi Chen, Jiachen Lu, Xiatian Zhu, Li ZhangCVPR 2023
- Sparsely Annotated Semantic Segmentation with Adaptive Gaussian MixturesLinshan Wu, Zhun Zhong, Leyuan Fang, Xingxin He 等CVPR 2023
- Open-World Instance Segmentation: Exploiting Pseudo Ground Truth From Learned Pairwise AffinityWeiyao Wang, Matt Feiszli, Heng Wang, Jitendra Malik 等CVPR 2022 · 被引用 39 次
- Rethinking Bayesian Deep Learning Methods for Semi-Supervised Volumetric Medical Image SegmentationJianfeng Wang, Thomas LukasiewiczCVPR 2022 · 被引用 31 次
- Generalize or Detect? Towards Robust Semantic Segmentation Under Multiple Distribution ShiftsZhitong Gao, Bingnan Li, Mathieu Salzmann, Xuming HeNeurIPS 2024 · 被引用 10 次
