Interpretable and Personalized Apprenticeship Scheduling: Learning Interpretable Scheduling Policies from Heterogeneous User Demonstrations
Rohan R. Paleja, Andrew Silva, Letian Chen, Matthew C. Gombolay
摘要
Resource scheduling and coordination is an NP-hard optimization requiring an efficient allocation of agents to a set of tasks with upper-and lower bound temporal and resource constraints. Due to the large-scale and dynamic nature of resource coordination in hospitals and factories, human domain experts manually plan and adjust schedules on the fly. To perform this job, domain experts leverage heterogeneous strategies and rules-of-thumb honed over years of apprenticeship. What is critically needed is the ability to extract this domain knowledge in a heterogeneous and interpretable apprenticeship learning framework to scale beyond the power of a single human expert, a necessity in safety-critical domains. We propose a personalized and interpretable apprenticeship scheduling algorithm that infers an interpretable representation of all human task demonstrators by extracting decision-making criteria via an inferred, personalized embedding non-parametric in the number of demonstrator types. We achieve near-perfect LfD accuracy in synthetic domains and 88.22% accuracy on a planning domain with real-world data, outperforming baselines. Finally, our user study showed our methodology produces more interpretable and easier-to-use models than neural networks (p < 0.05).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Encoding Human Domain Knowledge to Warm Start Reinforcement LearningAndrew Silva, Matthew C. GombolayAAAI 2021 · 被引用 42 次
- Fair Scheduling for Time-dependent ResourcesBo Li, Minming Li, Ruilong ZhangNeurIPS 2021 · 被引用 23 次
- Mixed-Initiative Multiagent Apprenticeship Learning for Human Training of Robot TeamsEsmaeil Seraj, Jerry Xiong, Mariah Schrum, Matthew C. GombolayNeurIPS 2023 · 被引用 12 次
- STL: Still Tricky Logic (for System Validation, Even When Showing Your Work)Isabelle Hurley, Rohan Paleja, Ashley Suh, Jaime Daniel Peña 等NeurIPS 2024 · 被引用 9 次
- Heterogeneous Graph Transformers for Simultaneous Mobile Multi-Robot Task Allocation and Scheduling under Temporal ConstraintsBatuhan Altundas, Shengkang Chen, Shivika Singh, Shivangi Deo 等NeurIPS 2025 · 被引用 1 次
相关 Paper
- Factorized Scheduling Principle: Learning Interpretable and Transferable Policies via Structured Additive FunctionsHong Je-Gal, Hyun-Suk LeeICML 2026
- Delphi: A Neuro-Symbolic Framework for Individualized, Safe and Interpretable Treatment RecommendationMuchan Tao, Haonan Qin, Yuqi Fang, Caifeng Shan 等AAAI 2026
- Hierarchical Causal Abduction: A Foundation Framework for Explainable Model Predictive ControlRamesh Arvind Naagarajan, Zühal Wagner, Stefan StreifICML 2026
- THEMES: An Offline Apprenticeship Learning Framework for Evolving Reward FunctionsXi Yang, Md. Mirajul Islam, Ge Gao, Min ChiKDD 2025
- Deep Bayesian Nonparametric Learning of Rules and Plans from Demonstrations with a Learned Automaton PriorBrandon Araki, Kiran Vodrahalli, Thomas Leech, Cristian Ioan Vasile 等AAAI 2020 · 被引用 8 次
