FreeSeg: Unified, Universal and Open-Vocabulary Image Segmentation
Jie Qin, Jie Wu, Pengxiang Yan, Ming Li, Yuxi Ren, Xuefeng Xiao, Yitong Wang, Rui Wang, Shilei Wen, Xin Pan, Xingang Wang
Abstract
Figure 1. We propose FreeSeg, a generic framework to accomplish Unified, Universal and Open-Vocabulary Image Segmentation. (a) FreeSeg optimizes an all-in-one network via one-shot training. (b) FreeSeg employs the same architecture and parameters to handle diverse segmentation tasks seamlessly in the inference procedure. (c) FreeSeg establishes new state-of-the-art performance across diverse segmentation tasks, training datasets and zero-shot generalization.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext be5d7c12-300f-47ec-a93e-4299b59aee4aCited by top-tier papers50
- Convolutions Die Hard: Open-Vocabulary Segmentation with Single Frozen Convolutional CLIPQihang Yu, Ju He, Xueqing Deng, Xiaohui Shen et al.NeurIPS 2023 · 285 citations
- Learning Mask-aware CLIP Representations for Zero-Shot SegmentationSiyu Jiao, Yunchao Wei, Yaowei Wang, Yao Zhao et al.NeurIPS 2023 · 88 citations
- Open-Vocabulary Video Anomaly DetectionPeng Wu, Xuerong Zhou, Guansong Pang, Yujia Sun et al.CVPR 2024 · 56 citations
- AlignDet: Aligning Pre-training and Fine-tuning in Object DetectionMing Li, Jie Wu, Xionghui Wang, Chen Chen et al.ICCV 2023 · 36 citations
- Relevant Intrinsic Feature Enhancement Network for Few-Shot Semantic SegmentationXiaoyi Bao, Jie Qin, Siyang Sun, Xingang Wang et al.AAAI 2024 · 34 citations
Builds on19
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Scaling Up Visual and Vision-Language Representation Learning With Noisy Text SupervisionChao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen et al.ICML 2021 · 5,401 citations
- Finetuned Language Models are Zero-Shot LearnersJason Wei, Maarten Bosma, Vincent Y. Zhao, Kelvin Guu et al.ICLR 2022 · 4,966 citations
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- Test-Time Training with Self-Supervision for Generalization under Distribution ShiftsYu Sun, Xiaolong Wang, Zhuang Liu, John Miller et al.ICML 2020 · 1,220 citations
Related papers
- SegGPT: Towards Segmenting Everything In ContextXinlong Wang, Xiaosong Zhang, Yue Cao, Wen Wang et al.ICCV 2023 · 188 citations
- Universal Segmentation at Arbitrary Granularity with Language InstructionYong Liu, Cairong Zhang, Yitong Wang, Jiahao Wang et al.CVPR 2024 · 15 citations
- OneFormer: One Transformer to Rule Universal Image SegmentationJitesh Jain, Jiachen Li, MangTik Chiu, Ali Hassani et al.CVPR 2023
- MasQCLIP for Open-Vocabulary Universal Image SegmentationXin Xu, Tianyi Xiong, Zheng Ding, Zhuowen TuICCV 2023 · 57 citations
- One-Prompt to Segment All Medical ImagesJunde Wu, Min XuCVPR 2024 · 32 citations
