Rethinking Query-based Transformer for Continual Image Segmentation
Yuchen Zhu, Cheng Shi, Dingyou Wang, Jiajin Tang, Zhengxuan Wei, Yu Wu, Guanbin Li, Sibei Yang
摘要
Class-incremental/Continual image segmentation (CIS) aims to train an image segmenter in stages, where the set of available categories differs at each stage. To leverage the built-in objectness of query-based transformers, which mitigates catastrophic forgetting of mask proposals, current methods often decouple mask generation from the continual learning process. This study, however, identifies two key issues with decoupled frameworks: loss of plasticity and heavy reliance on input data order. To address these, we conduct an in-depth investigation of the built-in objectness and find that highly aggregated image features provide a shortcut for queries to generate masks through simple feature alignment. Based on this, we propose SimCIS, a simple yet powerful baseline for CIS. Its core idea is to directly select image features for query assignment, ensuring "perfect alignment" to preserve objectness, while simultaneously allowing queries to select new classes to promote plasticity. To further combat catastrophic forgetting of categories, we introduce cross-stage consistency in selection and an innovative "visual query"-based replay mechanism. Experiments demonstrate that SimCIS consistently outperforms state-of-the-art methods across various segmentation tasks, settings, splits, and input data orders. All models and codes will be made publicly available at https://github.com/SooLab/SimCIS .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Vision Transformers Need More Than RegistersCheng Shi, Yizhou Yu, Sibei YangCVPR 2026 · 被引用 17 次
- Intervene-All-Paths: Unified Mitigation of LVLM Hallucinations across Alignment FormatsJiaye Qian, Ge Zheng, Yuchen Zhu, Sibei YangNeurIPS 2025 · 被引用 11 次
- Sim-DETR: Unlock DETR for Temporal Sentence GroundingJiajin Tang, Zhengxuan Wei, Yuchen Zhu, Cheng Shi 等ICCV 2025 · 被引用 3 次
- Why LVLMs are More Prone to Hallucinations in Longer Responses: The Role of ContextGe Zheng, Jiaye Qian, Jiajin Tang, Sibei YangICCV 2025 · 被引用 2 次
- Beyond Prompt Degradation: Prototype-guided Dual-pool Prompting for Incremental Object DetectionYaoteng Zhang, Qing Zhou, Junyu Gao, Qi WangCVPR 2026 · 被引用 2 次
它引用的顶会 Paper32
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang 等ICCV 2019 · 被引用 2,972 次
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
- DyTox: Transformers for Continual Learning with DYnamic TOken eXpansionArthur Douillard, Alexandre Ramé, Guillaume Couairon, Matthieu CordCVPR 2022 · 被引用 315 次
- DDCoT: Duty-Distinct Chain-of-Thought Prompting for Multimodal Reasoning in Language ModelsGe Zheng, Bin Yang, Jiajin Tang, Hong-Yu Zhou 等NeurIPS 2023 · 被引用 252 次
相关 Paper
- Continual Segmentation with Disentangled Objectness Learning and Class RecognitionYizheng Gong, Siyue Yu, Xiaoyang Wang, Jimin XiaoCVPR 2024
- Incrementer: Transformer for Class-Incremental Semantic Segmentation with Knowledge Distillation Focusing on Old ClassChao Shang, Hongliang Li, Fanman Meng, Qingbo Wu 等CVPR 2023
- Mining Unseen Classes via Regional Objectness: A Simple Baseline for Incremental SegmentationZekang Zhang, Guangyu Gao, Zhiyuan Fang, Jianbo Jiao 等NeurIPS 2022 · 被引用 53 次
- SSUL: Semantic Segmentation with Unknown Label for Exemplar-based Class-Incremental LearningSungmin Cha, Beomyoung Kim, Youngjoon Yoo, Taesup MoonNeurIPS 2021 · 被引用 139 次
- Beyond Background Shift: Rethinking Instance Replay in Continual Semantic SegmentationHongmei Yin, Tingliang Feng, Fan Lyu, Fanhua Shang 等CVPR 2025
