Lune

ICML2022顶会

Understanding Gradual Domain Adaptation: Improved Analysis, Optimal Path and Beyond

Haoxiang Wang, Bo Li, Han Zhao

2022年份
48被引次数
14顶会引用

摘要

The vast majority of existing algorithms for unsupervised domain adaptation (UDA) focus on adapting from a labeled source domain to an unlabeled target domain directly in a one-off way. Gradual domain adaptation (GDA), on the other hand, assumes a path of (T−1)(T-1) unlabeled intermediate domains bridging the source and target, and aims to provide better generalization in the target domain by leveraging the intermediate ones. Under certain assumptions, Kumar et al. (2020) proposed a simple algorithm, Gradual Self-Training, along with a generalization bound in the order of eO(T)(ε0+O(log(T)/n))e^{O(T)} \left(\varepsilon_0+O\left(\sqrt{log(T)/n}\right)\right) for the target domain error, where ε0\varepsilon_0 is the source domain error and nn is the data size of each domain. Due to the exponential factor, this upper bound becomes vacuous when TT is only moderately large. In this work, we analyze gradual self-training under more general and relaxed assumptions, and prove a significantly improved generalization bound as ε0+O(TΔ+T/n)+O~(1/nT)\varepsilon_0+ O \left(T\Delta + T/\sqrt{n}\right) + \widetilde{O}\left(1/\sqrt{nT}\right), where Δ\Delta is the average distributional distance between consecutive domains. Compared with the existing bound with an exponential dependency on TT as a multiplicative factor, our bound only depends on TT linearly and additively. Perhaps more interestingly, our result implies the existence of an optimal choice of TT that minimizes the generalization error, and it also naturally suggests an optimal way to construct the path of intermediate domains so as to minimize the accumulative path length TΔT\Delta between the source and target. To corroborate the implications of our theory, we examine gradual self-training on multiple semi-synthetic and real datasets, which confirms our findings. We believe our insights provide a path forward toward the design of future GDA algorithms.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper14

问问它们各自怎么用它

它引用的顶会 Paper11

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖