Self-Paced Video Data Augmentation by Generative Adversarial Networks with Insufficient Samples
Yumeng Zhang, Gaoguo Jia, Li Chen, Mingrui Zhang, Junhai Yong
摘要
An effective video classification method by means of a small number of samples is urgently needed. The deficiency of samples could be alleviated by generating samples through generative adversarial networks (GANs). However, the generation of videos in a typical category remains underexplored because the complex actions and the changeable viewpoints are difficult to simulate. Thus, applying GANs to perform video augmentation is difficult. In this study, we propose a generative data augmentation method for video classification using dynamic images. The dynamic image compresses the motion information of a video into a still image, removing the interference factors such as the background. Thus, utilizing the GANs to augment dynamic images can keep the categorical motion information and save memory compared with generating videos. To deal with the uneven quality of generated images, we propose a self-paced selection method to automatically select high-quality generated samples for training. These selected dynamic images are used to enhance the features, attain regularization, and finally achieve video augmentation. Our method is verified on two benchmark datasets, namely, HMDB51 and UCF101. Experimental results show that the method remarkably improves the accuracy of video classification under the circumstance of sample insufficiency and sample imbalance.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper3
- Similar Scenes Arouse Similar Emotions: Parallel Data Augmentation for Stylized Image CaptioningGuodun Li, Yuchen Zhai, Zehao Lin, Yin ZhangACM MM 2021 · 被引用 23 次
- Benchmarking the Robustness of Temporal Action Detection Models Against Temporal CorruptionsRunhao Zeng, Xiaoyong Chen, Jiaming Liang, Huisi Wu 等CVPR 2024 · 被引用 6 次
- Augmenting Moment Retrieval: Zero-Dependency Two-Stage LearningZhengxuan Wei, Jiajin Tang, Sibei YangICCV 2025
相关 Paper
- Exploring Temporally Dynamic Data Augmentation for Video RecognitionTaeoh Kim, Jinhyung Kim, Minho Shim, Sangdoo Yun 等ICLR 2023 · 被引用 5 次
- Time-Equivariant Contrastive Video Representation LearningSimon Jenni, Hailin JinICCV 2021 · 被引用 64 次
- A Slow-I-Fast-P Architecture for Compressed Video Action RecognitionJiapeng Li, Ping Wei, Yongchi Zhang, Nanning ZhengACM MM 2020 · 被引用 49 次
- DynamoNet: Dynamic Action and Motion NetworkAli Diba, Vivek Sharma, Luc Van Gool, Rainer StiefelhagenICCV 2019 · 被引用 123 次
- AWSD: Adaptive Weighted Spatiotemporal Distillation for Video RepresentationMohammad Tavakolian, Hamed Rezazadegan Tavakoli, Abdenour HadidICCV 2019 · 被引用 7 次
