Learning to Solve Complex Problems via Dataset Decomposition
Wanru Zhao, Lucas Page-Caccia, Zhengyan Shi, Minseon Kim, Weijia Xu, Alessandro Sordoni
摘要
Curriculum learning is a class of training strategies that organizes the data being exposed to a model by difficulty, gradually from simpler to more complex examples. This research explores a reverse curriculum generation approach that recursively decomposes complex datasets into simpler, more learnable components. We propose a teacher-student framework where the teacher is equipped with the ability to reason step-by-step, which is used to recursively generate easier versions of examples, enabling the student model to progressively master difficult tasks. We propose a novel scoring system to measure data difficulty based on its structural complexity and conceptual depth, allowing curriculum construction over decomposed data. Experiments on math datasets (MATH and AIME) and code generation datasets demonstrate that models trained with curricula generated by our approach exhibit superior performance compared to standard training on original datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- Let's Verify Step by StepHunter Lightman, Vineet Kosaraju, Yuri Burda, Harrison Edwards 等ICLR 2024 · 被引用 3,045 次
- STaR: Bootstrapping Reasoning With ReasoningEric Zelikman, Yuhuai Wu, Jesse Mu, Noah D. GoodmanNeurIPS 2022 · 被引用 1,126 次
- MetaMath: Bootstrap Your Own Mathematical Questions for Large Language ModelsLonghui Yu, Weisen Jiang, Han Shi, Jincheng Yu 等ICLR 2024 · 被引用 637 次
- When Do Curricula Work?Xiaoxia Wu, Ethan Dyer, Behnam NeyshaburICLR 2021 · 被引用 141 次
- Key-Point-Driven Data Synthesis with Its Enhancement on Mathematical ReasoningYiming Huang, Xiao Liu, Yeyun Gong, Zhibin Gou 等AAAI 2025 · 被引用 74 次
相关 Paper
- What Makes a Good Curriculum? Disentangling the Effects of Data Ordering on LLM Mathematical ReasoningYaning Jia, Chunhui Zhang, Xingjian Diao, Xiangchi Yuan 等ACL 2026 · 被引用 4 次
- In-sample Curriculum Learning by Sequence Completion for Natural Language GenerationQi Jia, Yizhu Liu, Haifeng Tang, Kenny Q. ZhuACL 2023 · 被引用 1 次
- Denoising Task Difficulty-based Curriculum for Training Diffusion ModelsJin-Young Kim, Hyojun Go, Soonwoo Kwon, Hyun-Gyoon KimICLR 2025
- Curriculum Reinforcement Learning via Constrained Optimal TransportPascal Klink, Haoyi Yang, Carlo D'Eramo, Jan Peters 等ICML 2022 · 被引用 44 次
- Curriculum Learning for Natural Language UnderstandingBenfeng Xu, Licheng Zhang, Zhendong Mao, Quan Wang 等ACL 2020 · 被引用 156 次
