Adversarial Bipartite Graph Learning for Video Domain Adaptation
Yadan Luo, Zi Huang, Zijian Wang, Zheng Zhang, Mahsa Baktashmotlagh
Abstract
Domain adaptation techniques, which focus on adapting models between distributionally different domains, are rarely explored in the video recognition area due to the significant spatial and temporal shifts across the source (i.e. training) and target (i.e. test) domains. As such, recent works on visual domain adaptation which leverage adversarial learning to unify the source and target video representations and strengthen the feature transferability are not highly effective on the videos. To overcome this limitation, in this paper, we learn a domain-agnostic video classifier instead of learning domain-invariant representations, and propose an Adversarial Bipartite Graph (ABG) learning framework which directly models the source-target interactions with a network topology of the bipartite graph. Specifically, the source and target frames are sampled as heterogeneous vertexes while the edges connecting two types of nodes measure the affinity among them. Through message-passing, each vertex aggregates the features from its heterogeneous neighbors, forcing the features coming from the same class to be mixed evenly. Explicitly exposing the video classifier to such cross-domain representations at the training and test stages makes our model less biased to the labeled source data, which in-turn results in achieving a better generalization on the target domain. The proposed framework is agnostic to the choices of frame aggregation, and therefore, four different aggregation functions are investigated for capturing appearance and temporal dynamics. To further enhance the model capacity and testify the robustness of the proposed architecture on difficult transfer tasks, we extend our model to work in a semi-supervised setting using an additional video-level bipartite graph. Extensive experiments conducted on four benchmark datasets evidence the effectiveness of the proposed approach over the state-of-the-art methods on the task of video recognition.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 74526fc8-f8b2-46ee-9259-da5ca6a11e65Cited by top-tier papers16
- Contrast and Mix: Temporal Contrastive Video Domain Adaptation with Background MixingAadarsh Sahoo, Rutav Shah, Rameswar Panda, Kate Saenko et al.NeurIPS 2021 · 89 citations
- Learning Bounds for Open-Set LearningZhen Fang, Jie Lu, Anjin Liu, Feng Liu et al.ICML 2021 · 67 citations
- Dual Bipartite Graph Learning: A General Approach for Domain Adaptive Object DetectionChaoqi Chen, Jiongcheng Li, Zebiao Zheng, Yue Huang et al.ICCV 2021 · 65 citations
- Revisiting Domain-Adaptive 3D Object Detection by Reliable, Diverse and Class-balanced Pseudo-LabelingZhuoxiao Chen, Yadan Luo, Zheng Wang, Mahsa Baktashmotlagh et al.ICCV 2023 · 40 citations
- Unsupervised Video Domain Adaptation for Action Recognition: A Disentanglement PerspectivePengfei Wei, Lingdong Kong, Xinghua Qu, Yi Ren et al.NeurIPS 2023 · 39 citations
Builds on4
- Temporal Attentive Alignment for Large-Scale Video Domain AdaptationMin-Hung Chen, Zsolt Kira, Ghassan Alregib, Jaekwon Yoo et al.ICCV 2019 · 205 citations
- Adversarial Cross-Domain Action Recognition with Co-AttentionBoxiao Pan, Zhangjie Cao, Ehsan Adeli, Juan Carlos NieblesAAAI 2020 · 114 citations
- Progressive Graph Learning for Open-Set Domain AdaptationYadan Luo, Zijian Wang, Zi Huang, Mahsa BaktashmotlaghICML 2020 · 114 citations
- Learning from the Past: Continual Meta-Learning with Bayesian Graph Neural NetworksYadan Luo, Zi Huang, Zheng Zhang, Ziwei Wang et al.AAAI 2020 · 27 citations
Related papers
- Spatial-temporal Causal Inference for Partial Image-to-video AdaptationJin Chen, Xinxiao Wu, Yao Hu, Jiebo LuoAAAI 2021 · 19 citations
- Unsupervised Domain Adaptive Graph Convolutional NetworksMan Wu, Shirui Pan, Chuan Zhou, Xiaojun Chang et al.WWW 2020 · 221 citations
- Partial Video Domain Adaptation with Partial Adversarial Temporal Attentive NetworkYuecong Xu, Jianfei Yang, Haozhi Cao, Zhenghua Chen et al.ICCV 2021 · 32 citations
- Discovering Informative and Robust Positives for Video Domain AdaptationChang Liu, Kunpeng Li, Michael Stopa, Jun Amano et al.ICLR 2023
- Diversifying Spatial-Temporal Perception for Video Domain GeneralizationKun-Yu Lin, Jia-Run Du, Yipeng Gao, Jiaming Zhou et al.NeurIPS 2023 · 27 citations
