Self-Paced Video Data Augmentation by Generative Adversarial Networks with Insufficient Samples
Yumeng Zhang, Gaoguo Jia, Li Chen, Mingrui Zhang, Junhai Yong
Abstract
An effective video classification method by means of a small number of samples is urgently needed. The deficiency of samples could be alleviated by generating samples through generative adversarial networks (GANs). However, the generation of videos in a typical category remains underexplored because the complex actions and the changeable viewpoints are difficult to simulate. Thus, applying GANs to perform video augmentation is difficult. In this study, we propose a generative data augmentation method for video classification using dynamic images. The dynamic image compresses the motion information of a video into a still image, removing the interference factors such as the background. Thus, utilizing the GANs to augment dynamic images can keep the categorical motion information and save memory compared with generating videos. To deal with the uneven quality of generated images, we propose a self-paced selection method to automatically select high-quality generated samples for training. These selected dynamic images are used to enhance the features, attain regularization, and finally achieve video augmentation. Our method is verified on two benchmark datasets, namely, HMDB51 and UCF101. Experimental results show that the method remarkably improves the accuracy of video classification under the circumstance of sample insufficiency and sample imbalance.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get cb90bdb7-1a19-4927-b5bd-545a1db5b9d4Cited by top-tier papers3
- Similar Scenes Arouse Similar Emotions: Parallel Data Augmentation for Stylized Image CaptioningGuodun Li, Yuchen Zhai, Zehao Lin, Yin ZhangACM MM 2021 · 23 citations
- Benchmarking the Robustness of Temporal Action Detection Models Against Temporal CorruptionsRunhao Zeng, Xiaoyong Chen, Jiaming Liang, Huisi Wu et al.CVPR 2024 · 6 citations
- Augmenting Moment Retrieval: Zero-Dependency Two-Stage LearningZhengxuan Wei, Jiajin Tang, Sibei YangICCV 2025
Related papers
- Exploring Temporally Dynamic Data Augmentation for Video RecognitionTaeoh Kim, Jinhyung Kim, Minho Shim, Sangdoo Yun et al.ICLR 2023 · 5 citations
- Time-Equivariant Contrastive Video Representation LearningSimon Jenni, Hailin JinICCV 2021 · 64 citations
- A Slow-I-Fast-P Architecture for Compressed Video Action RecognitionJiapeng Li, Ping Wei, Yongchi Zhang, Nanning ZhengACM MM 2020 · 49 citations
- DynamoNet: Dynamic Action and Motion NetworkAli Diba, Vivek Sharma, Luc Van Gool, Rainer StiefelhagenICCV 2019 · 123 citations
- AWSD: Adaptive Weighted Spatiotemporal Distillation for Video RepresentationMohammad Tavakolian, Hamed Rezazadegan Tavakoli, Abdenour HadidICCV 2019 · 7 citations
