OptMATH: A Scalable Bidirectional Data Synthesis Framework for Optimization Modeling
Hongliang Lu, Zhonglin Xie, Yaoyu Wu, Can Ren, Yuxuan Chen, Zaiwen Wen
摘要
Despite the rapid development of large language models (LLMs), a fundamental challenge persists: the lack of high-quality optimization modeling datasets hampers LLMs' robust modeling of practical optimization problems from natural language descriptions (NL). This data scarcity also contributes to the generalization difficulties experienced by learning-based methods. To address these challenges, we propose a scalable framework for synthesizing a high-quality dataset, named OptMATH. Starting from curated seed data with mathematical formulations (MF), this framework automatically generates problem data (PD) with controllable complexity. Then, a back-translation step is employed to obtain NL. To verify the correspondence between the NL and the PD, a forward modeling step followed by rejection sampling is used. The accepted pairs constitute the training part of OptMATH. Then a collection of rejected pairs is identified and further filtered. This collection serves as a new benchmark for optimization modeling, containing difficult instances whose lengths are much longer than these of NL4OPT and MAMO. Through extensive experiments, we demonstrate that models of various sizes (0.5B-32B parameters) trained on OptMATH achieve superior results on multiple modeling benchmarks, thereby validating the effectiveness and scalability of our ap-proach. Our dataset is publicly available at https://github.com/AuroraLHL/OptMATH .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Solver-Informed RL: Grounding Large Language Models for Authentic Optimization ModelingYitian Chen, Jingfan Xia, Siyu Shao, Dongdong Ge 等NeurIPS 2025 · 被引用 54 次
- OptiTree: Hierarchical Thoughts Generation with Tree Search for LLM Optimization ModelingHaoyang Liu, Jie Wang, Yuyang Cai, Xiongwei Han 等NeurIPS 2025 · 被引用 33 次
- StepORLM: A Self-Evolving Framework With Generative Process Supervision For Operations Research Language ModelsChenyu Zhou, Tianyi Xu, Jianghao Lin, Dongdong GeICLR 2026 · 被引用 29 次
- AlphaOPT: Formulating Optimization Programs with Self-Improving LLM Experience LibraryMinwei Kong, Ao Qu, Xiaotong Guo, Wenbin Ouyang 等KDD 2026 · 被引用 16 次
- MURKA: Multi-Reward Reinforcement Learning with Knowledge Alignment for Optimization TasksWantong Xie, Yi-Xiang Hu, Jieyang Xu, Feng Wu 等NeurIPS 2025 · 被引用 8 次
它引用的顶会 Paper12
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Self-Refine: Iterative Refinement with Self-FeedbackAman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan 等NeurIPS 2023 · 被引用 4,972 次
- Self-Instruct: Aligning Language Models with Self-Generated InstructionsYizhong Wang, Yeganeh Kordi, Swaroop Mishra, Alisa Liu 等ACL 2023 · 被引用 540 次
- Chain-of-Experts: When LLMs Meet Complex Operations Research ProblemsZiyang Xiao, Dongxiang Zhang, Yangjun Wu, Lilin Xu 等ICLR 2024 · 被引用 136 次
相关 Paper
- OptiBench Meets ReSocratic: Measure and Improve LLMs for Optimization ModelingZhicheng Yang, Yiwei Wang, Yinya Huang, Zhijiang Guo 等ICLR 2025
- ProOPF: Benchmarking and Improving LLMs for Professional-Grade Power Systems Optimization ModelingChao Shen, Zihan Guo, Xu Wan, Zhenghao Yang 等ICML 2026 · 被引用 3 次
- LLMOPT: Learning to Define and Solve General Optimization Problems from ScratchCaigao Jiang, Xiang Shu, Hong Qian, Xingyu Lu 等ICLR 2025
- MathSmith: Towards Extremely Hard Mathematical Reasoning by Forging Synthetic Problems with a Reinforced PolicyShaoxiong Zhan, Yanlin Lai, Ziyu Lu, Dahua Lin 等AAAI 2026 · 被引用 17 次
- HARDMath: A Benchmark Dataset for Challenging Problems in Applied MathematicsJingxuan Fan, Sarah Martinson, Erik Y. Wang, Kaylie Hausknecht 等ICLR 2025
