Making Learner Weakness Actionable for Learning from Demonstration with Novice Teachers
Yuqing Zhu, Matthew Howard
Abstract
Learning from demonstration can be an effective way to teach robots task-oriented policies. However, in an interactive setting when demonstrations are limited by time or other budgetary constraints, it is challenging to find those that fix the learner's (remaining) errors. This is especially difficult for novice teachers: they may provide task-valid trajectories, often these fail to meaningfully improve the policy due to their lack of knowledge of learning mechanisms internal to the robot. This paper introduces CLASP (Collaborative Learning with Anchored State-space Partitions), which summarises the teaching process as a compact map of behavioural regions anchored in the teacher's own demonstrations. The map connects task failure to actionable changes to demonstrations by indicating what is going wrong in an intuitive way. It also enables difficulty-aware training that emphasises regions where learning is failing. Across diverse benchmarks, CLASP improves success by up to 20% over offline and interactive baselines under the same demonstration budget, improves robustness under distribution shift by 14–20%, and preserves behavioural diversity.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e030a107-423f-447e-8fbe-6c4bb19daf45Builds on4
- Data Quality in Imitation LearningSuneel Belkhale, Yuchen Cui, Dorsa SadighNeurIPS 2023 · 135 citations
- Towards Diverse Behaviors: A Benchmark for Imitation Learning with Human DemonstrationsXiaogang Jia, Denis Blessing, Xinkai Jiang, Moritz Reuss et al.ICLR 2024 · 48 citations
- Reverse Forward Curriculum Learning for Extreme Sample and Demo EfficiencyStone Tao, Arth Shukla, Tse-kai Chan, Hao SuICLR 2024 · 6 citations
- Robot-Gated Interactive Imitation Learning with Adaptive Intervention MechanismHaoyuan Cai, Zhenghao Peng, Bolei ZhouICML 2025
Related papers
- STRAP: Robot Sub-Trajectory Retrieval for Augmented Policy LearningMarius Memmel, Jacob Berg, Bingqing Chen, Abhishek Gupta et al.ICLR 2025
- Active Fine-Tuning of Multi-Task PoliciesMarco Bagatella, Jonas Hübotter, Georg Martius, Andreas KrauseICML 2025
- Generalizable Coarse-to-Fine Robot Manipulation via Language-Aligned 3D KeypointsJianshu Hu, Lidi Wang, Shujia Li, Yunpeng Jiang et al.ICLR 2026 · 6 citations
- Robust Imitation of a Few Demonstrations with a Backwards ModelJung Yeon Park, Lawson L. S. WongNeurIPS 2022 · 21 citations
- KISA: A Unified Keyframe Identifier and Skill Annotator for Long-Horizon Robotics DemonstrationsLongxin Kou, Fei Ni, Yan Zheng, Jinyi Liu et al.ICML 2024 · 5 citations
