Teaching to Learn: Sequential Teaching of Learners with Internal States
Mustafa Mert Çelikok, Pierre-Alexandre Murena, Samuel Kaski
摘要
In sequential machine teaching, a teacher’s objective is to provide the optimal sequence of inputs to sequential learners in order to guide them towards the best model. However, this teaching objective considers a restricted class of learners with fixed inductive biases. In this paper, we extend the machine teaching framework to learners that can improve their inductive biases, represented as latent internal states, in order to generalize to new datasets. We introduce a novel framework in which learners’ inductive biases may change with the teaching interaction, which affects the learning performance in future tasks. In order to teach such learners, we propose a multi-objective control approach that takes the future performance of the learner after teaching into account. This framework provides tools for modelling learners with internal states, humans and meta-learning algorithms alike. Furthermore, we distinguish manipulative teaching, which can be done by effectively hiding data and also used for indoctrination, from teaching to learn which aims to help the learner become better at learning from new datasets in the absence of a teacher. Our empirical results demonstrate that our framework is able to reduce the number of required tasks for online meta-learning, and increases independent learning performance of simulated human users in future tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Synthesizing Programmatic Policies that Inductively GeneralizeJeevana Priya Inala, Osbert Bastani, Zenna Tavares, Armando Solar-LezamaICLR 2020 · 被引用 54 次
- Teaching Active Human LearnersZizhe Wang, Hailong SunAAAI 2021 · 被引用 1 次
- Teaching with Limited Information on the Learner's BehaviourFerdinando Cicalese, Sergio Filho, Eduardo Sany Laber, Marco MolinaroICML 2020 · 被引用 17 次
- On the Stability-Plasticity Dilemma in Continual Meta-Learning: Theory and AlgorithmQi Chen, Changjian Shui, Ligong Han, Mario MarchandNeurIPS 2023 · 被引用 32 次
- Learning to Reweight with Deep InteractionsYang Fan, Yingce Xia, Lijun Wu, Shufang Xie 等AAAI 2021 · 被引用 6 次
