Active Curriculum Refinement for Reinforcement Learning
Zhenya Liu, Yuxin Chen
2026年份
摘要
In many RL domains, environments are linked by prerequisite relations—e.g., difficulty-increasing edits or parameter increments—which induce a directed acyclic curriculum graph (DAG). In practice, this structure is often exploited only implicitly, yet it can yield clear gains in training. We introduce PATH, a curriculum learning framework that performs active learning on the curriculum graph. PATH first expands coverage by sampling diverse curriculum paths, then reallocates training toward regions that remain unmastered. Experiments show that PATH leverages the graph structure to achieve strong robustness and generalization across diverse environments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper16
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- Deep Reinforcement Learning at the Edge of the Statistical PrecipiceRishabh Agarwal, Max Schwarzer, Pablo Samuel Castro, Aaron C. Courville 等NeurIPS 2021 · 被引用 1,067 次
- Leveraging Procedural Generation to Benchmark Reinforcement LearningKarl Cobbe, Christopher Hesse, Jacob Hilton, John SchulmanICML 2020 · 被引用 685 次
- Emergent Complexity and Zero-shot Transfer via Unsupervised Environment DesignMichael Dennis, Natasha Jaques, Eugene Vinitsky, Alexandre M. Bayen 等NeurIPS 2020 · 被引用 362 次
- Prioritized Level ReplayMinqi Jiang, Edward Grefenstette, Tim RocktäschelICML 2021 · 被引用 211 次
相关 Paper
- Curriculum Reinforcement Learning via Constrained Optimal TransportPascal Klink, Haoyi Yang, Carlo D'Eramo, Jan Peters 等ICML 2022 · 被引用 44 次
- Item-Difficulty-Aware Learning Path Recommendation: From a Real Walking PerspectiveHaotian Zhang, Shuanghong Shen, Bihan Xu, Zhenya Huang 等KDD 2024 · 被引用 3 次
- Safety-Prioritizing Curricula for Constrained Reinforcement LearningCevahir Köprülü, Thiago D. Simão, Nils Jansen, Ufuk TopcuICLR 2025
- PORTAL: Automatic Curricula Generation for Multiagent Reinforcement LearningJizhou Wu, Jianye Hao, Tianpei Yang, Xiaotian Hao 等AAAI 2024 · 被引用 12 次
- EAT-C: Environment-Adversarial sub-Task Curriculum for Efficient Reinforcement LearningShuang Ao, Tianyi Zhou, Jing Jiang, Guodong Long 等ICML 2022 · 被引用 6 次
