Continual Unsupervised Generative Modelling via Online Optimal Transport
Fei Ye, Adrian G. Bors, Kun Zhang
Abstract
Lately, deep generative models have achieved excellent results after learning pre-defined and static data distribution. Meanwhile, their performance on continual learning suffers from degeneration, caused by catastrophic forgetting. In this paper, we study the unsupervised generative modelling in a more realistic continual learning scenario, where class and task information are absent during both training and inference learning phases. To implement this goal, the proposed memory approach consists of a temporary memory system, which stores data examples while a dynamic expansion memory system would gradually preserve those samples that are crucial for long-term memorization. A novel memory expansion mechanism is then proposed, by employing optimal transport distances between the statistics of memorized samples and each newly seen datum. This paper proposes the Sinkhorn-based Dual Dynamic Memory (SDDM) method, by considering Sinkhorn distance as an optimal transport measure, for evaluating the significance of the data to be stored in the memory buffer. The Sinkhorn transport algorithm leads to preserving a diversity of samples within a compact memory capacity. The memory buffering approach does not interact with the model's training process and can be optimized independently in both supervised and unsupervised learning without any modifications. Moreover, we also propose a novel dynamic model expansion mechanism to automatically increase the model's capacity whenever necessary, which can deal with infinite data streams and further improve the model's performance. Experimental results show that the proposed approach achieves state-of-the-art performance in both supervised and unsupervised learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 575417c2-2a49-459d-b3aa-2d598de86f67Builds on20
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 5,234 citations
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- Continual Prototype Evolution: Learning Online from Non-Stationary Data StreamsMatthias De Lange, Tinne TuytelaarsICCV 2021 · 251 citations
- A Neural Dirichlet Process Mixture Model for Task-Free Continual LearningSoochan Lee, Junsoo Ha, Dongsu Zhang, Gunhee KimICLR 2020 · 238 citations
Related papers
- Online Task-Free Continual Learning via Dynamic Expansionable Memory DistributionFei Ye, Adrian G. BorsCVPR 2025
- Dynamic Expansion Diffusion Learning for Lifelong Generative ModellingFei Ye, Adrian G. Bors, Kun ZhangAAAI 2025 · 4 citations
- Online Task-Free Continual Generative and Discriminative Learning via Dynamic Cluster MemoryFei Ye, Adrian G. BorsCVPR 2024
- Lifelong Scalable Generative System via Online Maximum Mean DiscrepancyFei Ye, Adrian G. BorsAAAI 2025
- Task-Free Continual Generation and Representation Learning via Dynamic Expansionable Memory ClusterFei Ye, Adrian G. BorsAAAI 2024 · 8 citations
