When is Transfer Learning Possible?
My Phan, Kianté Brantley, Stephanie Milani, Soroush Mehri, Gokul Swamy, Geoffrey J. Gordon
Abstract
We present a general framework for transfer learning that is flexible enough to capture transfer in supervised, reinforcement, and imitation learning. Our framework enables new insights into the fundamental question of when we can successfully transfer learned information across problems. We model the learner as interacting with a sequence of problem instances, or environments, each of which is generated from a common structural causal model (SCM) by choosing the SCM's parameters from restricted sets. We derive a procedure that can propagate restrictions on SCM parameters through the SCM's graph structure to other parameters that we are trying to learn. The propagated restrictions then enable more efficient learning (i.e., transfer). By analyzing the procedure, we are able to challenge widely-held beliefs about transfer learning. First, we show that having sparse changes across environments is neither necessary nor sufficient for transfer. Second, we show an example where the common heuristic of freezing a layer in a network causes poor transfer performance. We then use our procedure to select a more refined set of parameters to freeze, leading to successful transfer learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 57e7c243-0498-4fa5-808a-d36072f76b6dCited by top-tier papers1
Ask how each one uses itBuilds on10
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 1,416 citations
- Bellman Eluder Dimension: New Rich Classes of RL Problems, and Sample-Efficient AlgorithmsChi Jin, Qinghua Liu, Sobhan MiryoosefiNeurIPS 2021 · 264 citations
- Reverse-engineering deep ReLU networksDavid Rolnick, Konrad P. KordingICML 2020 · 121 citations
- Connect, Not Collapse: Explaining Contrastive Learning for Unsupervised Domain AdaptationKendrick Shen, Robbie M. Jones, Ananya Kumar, Sang Michael Xie et al.ICML 2022 · 102 citations
- A Calculus for Stochastic Interventions: Causal Effect Identification and Surrogate ExperimentsJuan D. Correa, Elias BareinboimAAAI 2020 · 90 citations
Related papers
- Theory-Based Causal Transfer: Integrating Instance-Level Induction and Abstract-Level Structure LearningMark Edmonds, Xiaojian Ma, Siyuan Qi, Yixin Zhu et al.AAAI 2020 · 28 citations
- Few-shot Domain Adaptation by Causal Mechanism TransferTakeshi Teshima, Issei Sato, Masashi SugiyamaICML 2020 · 101 citations
- Comparing Causal Frameworks: Potential Outcomes, Structural Models, Graphs, and AbstractionsDuligur Ibeling, Thomas IcardNeurIPS 2023 · 27 citations
- Counterfactual Structural Causal BanditsMin Woo Park, Sanghack LeeICLR 2026 · 1 citation
- Synthesizing Programmatic Policies that Inductively GeneralizeJeevana Priya Inala, Osbert Bastani, Zenna Tavares, Armando Solar-LezamaICLR 2020 · 54 citations
