Calibrating Class Weights with Multi-Modal Information for Partial Video Domain Adaptation
Xiyu Wang, Yuecong Xu, Jianfei Yang, Kezhi Mao
Abstract
Assuming the source label space subsumes the target one, Partial Video Domain Adaptation (PVDA) is a more general and practical scenario for cross-domain video classification problems. The key challenge of PVDA is to mitigate the negative transfer caused by the source-only outlier classes. To tackle this challenge, a crucial step is to aggregate target predictions to assign class weights by up-weighing target classes and down-weighing outlier classes. However, the incorrect predictions of class weights can mislead the network and lead to negative transfer. Previous works improve the class weight accuracy by utilizing temporal features and attention mechanisms, but these methods may fall short when trying to generate accurate class weight when domain shifts are significant, as in most real-world scenarios. To deal with these challenges, we first propose the Multi-modality partial Adversarial Network (MAN), which utilizes multi-scale and multi-modal information to enhance PVDA performance. Based on MAN, we then propose Multi-modality Cluster-calibrated partial Adversarial Network (MCAN). It utilizes a novel class weight calibration method to alleviate the negative transfer caused by incorrect class weights. Specifically, the calibration method tries to identify and weigh correct and incorrect predictions using distributional information implied by unsupervised clustering. Extensive experiments are conducted on prevailing PVDA benchmarks, and the proposed MCAN achieves significant improvements when compared to state-of-the-art PVDA methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9d8bcc41-2c7c-47a7-932f-05d0ccf57e8eCited by top-tier papers3
- SWL-Adapt: An Unsupervised Domain Adaptation Model with Sample Weight Learning for Cross-User Wearable Human Activity RecognitionRong Hu, Ling Chen, Shenghuan Miao, Xing TangAAAI 2023 · 48 citations
- Can We Evaluate Domain Adaptation Models Without Target-Domain Labels?Jianfei Yang, Hanjie Qian, Yuecong Xu, Kai Wang et al.ICLR 2024 · 16 citations
- Augmenting and Aligning Snippets for Few-Shot Video Domain AdaptationYuecong Xu, Jianfei Yang, Yunjiao Zhou, Zhenghua Chen et al.ICCV 2023 · 10 citations
Builds on11
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 4,104 citations
- Do We Really Need to Access the Source Data? Source Hypothesis Transfer for Unsupervised Domain AdaptationJian Liang, Dapeng Hu, Jiashi FengICML 2020 · 1,624 citations
- Learning to Estimate Hidden Motions with Global Motion AggregationShihao Jiang, Dylan Campbell, Yao Lu, Hongdong Li et al.ICCV 2021 · 402 citations
- Cluster Alignment With a Teacher for Unsupervised Domain AdaptationZhijie Deng, Yucen Luo, Jun ZhuICCV 2019 · 241 citations
- Temporal Attentive Alignment for Large-Scale Video Domain AdaptationMin-Hung Chen, Zsolt Kira, Ghassan Alregib, Jaekwon Yoo et al.ICCV 2019 · 205 citations
Related papers
- Partial Video Domain Adaptation with Partial Adversarial Temporal Attentive NetworkYuecong Xu, Jianfei Yang, Haozhi Cao, Zhenghua Chen et al.ICCV 2021 · 32 citations
- Adversarial Reweighting for Partial Domain AdaptationXiang Gu, Xi Yu, Yan Yang, Jian Sun et al.NeurIPS 2021 · 60 citations
- Image-to-video Adaptation with Outlier Modeling and Robust Self-learningJunbao Zhuo, Shuhui Wang, Zhenghan Chen, Li Shen et al.AAAI 2025
- Mix-DANN and Dynamic-Modal-Distillation for Video Domain AdaptationYuehao Yin, Bin Zhu, Jingjing Chen, Lechao Cheng et al.ACM MM 2022 · 7 citations
- mDALU: Multi-Source Domain Adaptation and Label Unification with Partial DatasetsRui Gong, Dengxin Dai, Yuhua Chen, Wen Li et al.ICCV 2021 · 27 citations
