PIVOT: Prompting for Video Continual Learning
Andrés Villa, Juan León Alcázar, Motasem Alfarra, Kumail Alhamoud, Julio Hurtado, Fabian Caba Heilbron, Alvaro Soto, Bernard Ghanem
摘要
Modern machine learning pipelines are limited due to data availability, storage quotas, privacy regulations, and expensive annotation processes. These constraints make it difficult or impossible to train and update large-scale models on such dynamic annotated sets. Continual learning directly approaches this problem, with the ultimate goal of devising methods where a deep neural network effectively learns relevant patterns for new (unseen) classes, without significantly altering its performance on previously learned ones. In this paper, we address the problem of continual learning for video data. We introduce PIVOT, a novel method that leverages extensive knowledge in pre-trained models from the image domain, thereby reducing the number of trainable parameters and the associated forgetting. Unlike previous methods, ours is the first approach that effectively uses prompting mechanisms for continual learning without any in-domain pre-training. Our experiments show that PIVOT improves state-of-the-art methods by a significant 27% on the 20-task ActivityNet setup.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Generating Instance-level Prompts for Rehearsal-free Continual LearningDahuin Jung, Dongyoon Han, Jihwan Bang, Hwanjun SongICCV 2023 · 被引用 91 次
- Bisecle: Binding and Separation in Continual Learning for Video Language UnderstandingYue Tan, Xiaoqian Hu, Hao Xue, Celso de Melo 等NeurIPS 2025 · 被引用 14 次
- Continual Text-to-Video Retrieval with Frame Fusion and Task-Aware RoutingZecheng Zhao, Zhi Chen, Zi Huang, Shazia Sadiq 等SIGIR 2025 · 被引用 6 次
- Progressive Fourier Neural Representation for Sequential Video CompilationHaeyong Kang, Jaehong Yoon, Dahyun Kim, Sung Ju Hwang 等ICLR 2024 · 被引用 4 次
- RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual LearningKiseong Hong, Gyeong-Hyeon Kim, Eunwoo KimICCV 2025 · 被引用 3 次
它引用的顶会 Paper13
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Learning to Prompt for Continual LearningZifeng Wang, Zizhao Zhang, Chen-Yu Lee, Han Zhang 等CVPR 2022 · 被引用 635 次
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 被引用 409 次
- VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text UnderstandingHu Xu, Gargi Ghosh, Po-Yao Huang, Dmytro Okhonko 等EMNLP 2021 · 被引用 399 次
相关 Paper
- Label-Efficient Online Continual Object Detection in Streaming VideoJay Zhangjie Wu, David Junhao Zhang, Wynne Hsu, Mengmi Zhang 等ICCV 2023 · 被引用 24 次
- vCLIMB: A Novel Video Class Incremental Learning BenchmarkAndrés Villa, Kumail Alhamoud, Victor Escorcia, Fabian Caba Heilbron 等CVPR 2022 · 被引用 32 次
- Class-Incremental Learning for Action Recognition in VideosJaeyoo Park, Minsoo Kang, Bohyung HanICCV 2021 · 被引用 68 次
- DyStaB: Unsupervised Object Segmentation via Dynamic-Static BootstrappingYanchao Yang, Brian Lai, Stefano SoattoCVPR 2021
- AttriCLIP: A Non-Incremental Learner for Incremental Knowledge LearningRunqi Wang, Xiaoyue Duan, Guoliang Kang, Jianzhuang Liu 等CVPR 2023
