Adversarial Alignment with Anchor Dragging Drift (A³D²): Multimodal Domain Adaptation with Partially Shifted Modalities
Jun Sun, Xinxin Zhang, Simin Hong, Jian Zhu, Lingfang Zeng
摘要
Multimodal learning has celebrated remarkable success across diverse areas, yet faces the challenge of prohibitively expensive data collection and annotation when adapting models to new environments. In this context, domain adaptation has gained growing popularity as a technique for knowledge transfer, which, however, remains underexplored in multimodal settings compared with unimodal ones. This paper investigates multimodal domain adaptation, focusing on a practical partially shifting scenario where some modalities (referred to as anchors) remain domain-stable, while others (referred to as drifts) undergo a domain shift. We propose a bi-alignment scheme to simultaneously perform drift-drift and anchordrift matching. The former is achieved through adversarial learning, aligning the representations of the drifts across source and target domains; the latter corresponds to an "anchor dragging drift" strategy, which matches the distributions of the drifts and anchors within the target domain using the optimal transport (OT) method. The overall design principle features Adversarial Alignment with Anchor Dragging Drift, abbreviated as A 3 D 2 , for multimodal domain adaptation with partially shifted modalities. Comprehensive empirical results verify the effectiveness of the proposed approach, and demonstrate that A 3 D 2 achieves superior performance compared with state-of-the-art approaches. The code is available at: https: //github.com/sunjunaimer/A3D2.git .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper18
- Unbalanced minibatch Optimal Transport; applications to Domain AdaptationKilian Fatras, Thibault Séjourné, Rémi Flamary, Nicolas CourtyICML 2021 · 被引用 183 次
- How Does Information Bottleneck Help Deep Learning?Kenji Kawaguchi, Zhun Deng, Xu Ji, Jiaoyang HuangICML 2023 · 被引用 117 次
- Learning Cross-Modal Contrastive Features for Video Domain AdaptationDonghyun Kim, Yi-Hsuan Tsai, Bingbing Zhuang, Xiang Yu 等ICCV 2021 · 被引用 88 次
- MIntRec: A New Dataset for Multimodal Intent RecognitionHanlei Zhang, Hua Xu, Xin Wang, Qianrui Zhou 等ACM MM 2022 · 被引用 66 次
- Improving Mini-batch Optimal Transport via Partial TransportationKhai Nguyen, Dang Nguyen, The-Anh Vu-Le, Tung Pham 等ICML 2022 · 被引用 60 次
相关 Paper
- Boomda: Balanced Multi-objective Optimization for Multimodal Domain AdaptationJun Sun, Xinxin Zhang, Simin Hong, Jian Zhu 等AAAI 2026 · 被引用 1 次
- mDALU: Multi-Source Domain Adaptation and Label Unification with Partial DatasetsRui Gong, Dengxin Dai, Yuhua Chen, Wen Li 等ICCV 2021 · 被引用 27 次
- Multi-Anchor Active Domain Adaptation for Semantic SegmentationMunan Ning, Donghuan Lu, Dong Wei, Cheng Bian 等ICCV 2021 · 被引用 68 次
- Prototypical Partial Optimal Transport for Universal Domain AdaptationYucheng Yang, Xiang Gu, Jian SunAAAI 2023 · 被引用 21 次
- Uncertainty-Aware Alignment Network for Cross-Domain Video-Text RetrievalXiaoshuai Hao, Wanqian ZhangNeurIPS 2023 · 被引用 26 次
