Unsupervised Domain Adaptation for Video Object Grounding with Cascaded Debiasing Learning
Mengze Li, Haoyu Zhang, Juncheng Li, Zhou Zhao, Wenqiao Zhang, Shengyu Zhang, Shiliang Pu, Yueting Zhuang, Fei Wu
摘要
This paper addresses the Unsupervised Domain Adaptation (UDA) for the dense frame prediction task - Video Object Grounding (VOG). This investigation springs from the recognition of the limited generalization capabilities of data-driven approaches when confronted with unseen test scenarios. We set the goal of enhancing the adaptability of the source-dominated model from a labeled domain to the unlabeled target domain through re-training on pseudo-labels (i.e., predicted boxes of language-described objects). Given the potential for source-domain biases in the pseudo-label generation, we decompose the labeling refinement as two cascaded debiasing subroutines: (1) we develop a discarded training strategy to correct the Biased Proposal Selection by filtering out the examples with uncertain proposals selected from the proposal (candidate box) set. The identifier of these uncertain examples is the discordance between the predictions of the source-dominated model and those of a target-domain clustered classifier, which remains free from the source-domain bias. (2) With the refined proposals as a foundation, we measure Grounding Coordinate Offset based on the semantic distance of the model's prediction across domains, based on which we alleviate source-domain bias in the target model through adversarial learning. To verify the superiority of the proposed method, we collected two UDA-VOG datasets called I2O-VOG and R2M-VOG by manually dividing and combining the well-known VOG datasets. The extensive experiments on them show our model significantly outperforms SOTA methods by a large margin.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper5
- Fine-tuning Multimodal LLMs to Follow Zero-shot Demonstrative InstructionsJuncheng Li, Kaihang Pan, Zhiqi Ge, Minghe Gao 等ICLR 2024 · 被引用 95 次
- A Unified Approach to Domain Incremental Learning with Memory: Theory and AlgorithmHaizhou Shi, Hao WangNeurIPS 2023 · 被引用 60 次
- Revisiting the Domain Shift and Sample Uncertainty in Multi-source Active Domain TransferWenqiao Zhang, Zheqi LvCVPR 2024 · 被引用 17 次
- Unsupervised Domain Adaptation for Anatomical Structure Detection in Ultrasound ImagesBin Pu, Xingguo Lv, Jiewen Yang, Guannan He 等ICML 2024 · 被引用 10 次
- Leveraging Anatomical Consistency for Multi-Object Detection in Ultrasound Images via Source-free Unsupervised Domain AdaptationBin Pu, Xingguo Lv, Jiewen Yang, Xingbo Dong 等AAAI 2025 · 被引用 6 次
相关 Paper
- Category Dictionary Guided Unsupervised Domain Adaptation for Object DetectionShuai Li, Jianqiang Huang, Xian-Sheng Hua, Lei ZhangAAAI 2021 · 被引用 47 次
- Pseudo Label Refinery for Unsupervised Domain Adaptation on Cross-Dataset 3D Object DetectionZhanwei Zhang, Minghao Chen, Shuai Xiao, Liang Peng 等CVPR 2024 · 被引用 10 次
- Unsupervised Video Domain Adaptation with Masked Pre-Training and Collaborative Self-TrainingArun V. Reddy, William Paul, Corban Rivera, Ketul Shah 等CVPR 2024 · 被引用 3 次
- Hierarchical Debiasing and Noisy Correction for Cross-domain Video Tube RetrievalJingqiao Xiu, Mengze Li, Wei Ji, Jingyuan Chen 等ACM MM 2024 · 被引用 5 次
- IGG: Improved Graph Generation for Domain Adaptive Object DetectionPengteng Li, Ying He, F. Richard Yu, Pinhao Song 等ACM MM 2023 · 被引用 10 次
