Interpretable and Personalized Apprenticeship Scheduling: Learning Interpretable Scheduling Policies from Heterogeneous User Demonstrations
Rohan R. Paleja, Andrew Silva, Letian Chen, Matthew C. Gombolay
Abstract
Resource scheduling and coordination is an NP-hard optimization requiring an efficient allocation of agents to a set of tasks with upper-and lower bound temporal and resource constraints. Due to the large-scale and dynamic nature of resource coordination in hospitals and factories, human domain experts manually plan and adjust schedules on the fly. To perform this job, domain experts leverage heterogeneous strategies and rules-of-thumb honed over years of apprenticeship. What is critically needed is the ability to extract this domain knowledge in a heterogeneous and interpretable apprenticeship learning framework to scale beyond the power of a single human expert, a necessity in safety-critical domains. We propose a personalized and interpretable apprenticeship scheduling algorithm that infers an interpretable representation of all human task demonstrators by extracting decision-making criteria via an inferred, personalized embedding non-parametric in the number of demonstrator types. We achieve near-perfect LfD accuracy in synthetic domains and 88.22% accuracy on a planning domain with real-world data, outperforming baselines. Finally, our user study showed our methodology produces more interpretable and easier-to-use models than neural networks (p < 0.05).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d1d552e9-3e00-4a65-aeff-da06fc5da4dbCited by top-tier papers6
- Encoding Human Domain Knowledge to Warm Start Reinforcement LearningAndrew Silva, Matthew C. GombolayAAAI 2021 · 42 citations
- Fair Scheduling for Time-dependent ResourcesBo Li, Minming Li, Ruilong ZhangNeurIPS 2021 · 23 citations
- Mixed-Initiative Multiagent Apprenticeship Learning for Human Training of Robot TeamsEsmaeil Seraj, Jerry Xiong, Mariah Schrum, Matthew C. GombolayNeurIPS 2023 · 12 citations
- STL: Still Tricky Logic (for System Validation, Even When Showing Your Work)Isabelle Hurley, Rohan Paleja, Ashley Suh, Jaime Daniel Peña et al.NeurIPS 2024 · 9 citations
- Heterogeneous Graph Transformers for Simultaneous Mobile Multi-Robot Task Allocation and Scheduling under Temporal ConstraintsBatuhan Altundas, Shengkang Chen, Shivika Singh, Shivangi Deo et al.NeurIPS 2025 · 1 citation
Related papers
- Factorized Scheduling Principle: Learning Interpretable and Transferable Policies via Structured Additive FunctionsHong Je-Gal, Hyun-Suk LeeICML 2026
- Delphi: A Neuro-Symbolic Framework for Individualized, Safe and Interpretable Treatment RecommendationMuchan Tao, Haonan Qin, Yuqi Fang, Caifeng Shan et al.AAAI 2026
- Hierarchical Causal Abduction: A Foundation Framework for Explainable Model Predictive ControlRamesh Arvind Naagarajan, Zühal Wagner, Stefan StreifICML 2026
- THEMES: An Offline Apprenticeship Learning Framework for Evolving Reward FunctionsXi Yang, Md. Mirajul Islam, Ge Gao, Min ChiKDD 2025
- Deep Bayesian Nonparametric Learning of Rules and Plans from Demonstrations with a Learned Automaton PriorBrandon Araki, Kiran Vodrahalli, Thomas Leech, Cristian Ioan Vasile et al.AAAI 2020 · 8 citations
