Dual-Space Semantic Synergy Distillation for Continual Learning of Unlabeled Streams
Donghao Sun, Xi Wang, Xu Yang, Kun Wei, Cheng Deng
摘要
Continual learning from unlabeled data streams while effectively combating catastrophic forgetting poses an intractable challenge. Traditional methods predominantly rely on visual clustering techniques to generate pseudo labels, which often suffer from semantic inconsistencies and limited discriminative precision, thereby impeding stable model evolution. To surmount these obstacles, we introduce an innovative approach that synergistically combines both visual and textual information to generate dual space hybrid pseudo labels for reliable model continual evolution. Specifically, by harnessing the capabilities of large multimodal models, we initially generate generalizable text descriptions for a few representative samples. These descriptions then undergo a 'Coarse to Fine' refinement process to capture the subtle nuances between different data points, significantly enhancing the semantic accuracy of the descriptions. Simultaneously, a novel cross-modal hybrid approach seamlessly integrates these fine-grained textual descriptions with visual features, thereby creating a more robust and reliable supervisory signal. Finally, such descriptions are employed to alleviate the catastrophic forgetting issue via a semantic alignment distillation, which capitalizes on the stability inherent in language knowledge to effectively prevent the model from forgetting previously learned information. Comprehensive experiments conducted on a variety of benchmarks demonstrate that our proposed method attains state-of-the-art performance, and ablation studies further substantiate the effectiveness and superiority of the proposed method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang 等NeurIPS 2020 · 被引用 5,129 次
- In Defense of Pseudo-Labeling: An Uncertainty-Aware Pseudo-label Selection Framework for Semi-Supervised LearningMamshad Nayeem Rizve, Kevin Duarte, Yogesh S. Rawat, Mubarak ShahICLR 2021 · 被引用 630 次
- Learning to Discover Novel Visual Categories via Deep Transfer ClusteringKai Han, Andrea Vedaldi, Andrew ZissermanICCV 2019 · 被引用 378 次
- Improving CLIP Training with Language RewritesLijie Fan, Dilip Krishnan, Phillip Isola, Dina Katabi 等NeurIPS 2023 · 被引用 308 次
相关 Paper
- Knowledge Graph Enhanced Generative Multi-modal Models for Class-Incremental LearningXusheng Cao, Haori Lu, Linlan Huang, Fei Yang 等NeurIPS 2025 · 被引用 3 次
- Towards Dynamic Modality Alignment in Multimodal Continual LearningJiayao Tan, Fan Lyu, Tianle Liu, Fuyuan Hu 等CVPR 2026
- RECALL: REpresentation-aligned Catastrophic-forgetting ALLeviation via Hierarchical Model MergingBowen Wang, Haiyuan Wan, Liwen Shi, Chen Yang 等EMNLP 2025
- Federated Continual Learning via Orchestrating Multi-Scale ExpertiseXiaoyang Yi, Yang Liu, Binhan Yang, Jian Jun ZhangNeurIPS 2025
- Multi-Level Cross-Modal Alignment for Image ClusteringLiping Qiu, Qin Zhang, Xiaojun Chen, Shaotian CaiAAAI 2024 · 被引用 8 次
