STAND: Self-Aware Precondition Induction for Interactive Task Learning
Daniel Weitekamp, Glen Smith, Ken Koedinger, Christopher MacLellan
摘要
In interactive task learning (ITL), AI agents learn new capabilities from limited human instruction provided during task execution. STAND is a new method of data-efficient rule precondition induction specifically designed for these human-in-the-loop training scenarios. A key feature of STAND is its self-awareness of its own learning—it can provide accurate metrics of training progress back to users. STAND beats popular methods like XGBoost, decision trees, random forests, and version spaces at small-data precondition induction tasks, and is highly accurate at estimating when its performance improves on holdout examples. In our evaluations, we find that STAND shows more monotonic improvement than other models with low rates of error reoccurrence. These features of STAND support a consistent training experience, enabling human instructors to estimate when they have finished training and providing active-learning support by identifying trouble spots that require more training. STAND achieves this by efficiently learning a compact space of greedy classifiers consistent with training data, rather than a finite ensemble of alternatives.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- An Interaction Design for Machine Teaching to Develop AI TutorsDaniel Weitekamp III, Erik Harpstead, Kenneth R. KoedingerCHI 2020 · 被引用 69 次
- Hierarchical Shrinkage: Improving the accuracy and interpretability of tree-based modelsAbhineet Agarwal, Yan Shuo Tan, Omer Ronen, Chandan Singh 等ICML 2022 · 被引用 37 次
- VAL: Interactive Task Learning with GPT Dialog ParsingLane Lawley, Christopher MacLellanCHI 2024 · 被引用 15 次
相关 Paper
- PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-trainingKimin Lee, Laura M. Smith, Pieter AbbeelICML 2021 · 被引用 380 次
- Robot-Gated Interactive Imitation Learning with Adaptive Intervention MechanismHaoyuan Cai, Zhenghao Peng, Bolei ZhouICML 2025
- Teachable Reinforcement Learning via Advice DistillationOlivia Watkins, Abhishek Gupta, Trevor Darrell, Pieter Abbeel 等NeurIPS 2021 · 被引用 3 次
- Prompt-Based Rule Discovery and Boosting for Interactive Weakly-Supervised LearningRongzhi Zhang, Yue Yu, Pranav Shetty, Le Song 等ACL 2022 · 被引用 29 次
- Provable Interactive Learning with Hindsight Instruction FeedbackDipendra Misra, Aldo Pacchiano, Robert E. SchapireICML 2024 · 被引用 1 次
