Exploiting Motion Information from Unlabeled Videos for Static Image Action Recognition
Yiyi Zhang, Li Niu, Ziqi Pan, Meichao Luo, Jianfu Zhang, Dawei Cheng, Liqing Zhang
摘要
Static image action recognition, which aims to recognize action based on a single image, usually relies on expensive human labeling effort such as adequate labeled action images and large-scale labeled image dataset. In contrast, abundant unlabeled videos can be economically obtained. Therefore, several works have explored using unlabeled videos to facilitate image action recognition, which can be categorized into the following two groups: (a) enhance visual representations of action images with a designed proxy task on unlabeled videos, which falls into the scope of self-supervised learning; (b) generate auxiliary representations for action images with the generator learned from unlabeled videos. In this paper, we integrate the above two strategies in a unified framework, which consists of Visual Representation Enhancement (VRE) module and Motion Representation Augmentation (MRA) module. Specifically, the VRE module includes a proxy task which imposes pseudo motion label constraint and temporal coherence constraint on unlabeled videos, while the MRA module could predict the motion information of a static action image by exploiting unlabeled videos. We demonstrate the superiority of our framework based on four benchmark human action datasets with limited labeled data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Unsupervised Motion Representation Learning with Capsule AutoencodersZiwei Xu, Xudong Shen, Yongkang Wong, Mohan S. KankanhalliNeurIPS 2021 · 被引用 33 次
- Amodal Instance Segmentation via Prior-Guided ExpansionJunjie Chen, Li Niu, Jianfu Zhang, Jianlou Si 等AAAI 2023 · 被引用 25 次
- OSAN: A One-Stage Alignment Network to Unify Multimodal Alignment and Unsupervised Domain AdaptationYe Liu, Lingfeng Qiao, Changchong Lu, Di Yin 等CVPR 2023
相关 Paper
- Self-Supervised Motion Learning From Static ImagesZiyuan Huang, Shiwei Zhang, Jianwen Jiang, Mingqian Tang 等CVPR 2021
- Multiview Pseudo-Labeling for Semi-supervised Learning from VideoBo Xiong, Haoqi Fan, Kristen Grauman, Christoph FeichtenhoferICCV 2021 · 被引用 54 次
- RSPNet: Relative Speed Perception for Unsupervised Video Representation LearningPeihao Chen, Deng Huang, Dongliang He, Xiang Long 等AAAI 2021 · 被引用 140 次
- Exploiting Self-Supervised and Semi-Supervised Learning for Facial Landmark Tracking with Unlabeled DataShi Yin, Shangfei Wang, Xiaoping Chen, Enhong ChenACM MM 2020 · 被引用 7 次
- Unsupervised Video Domain Adaptation with Masked Pre-Training and Collaborative Self-TrainingArun V. Reddy, William Paul, Corban Rivera, Ketul Shah 等CVPR 2024 · 被引用 3 次
