Automated Proof of Polynomial Inequalities via Reinforcement Learning
Banglong Liu, Niuniu Qi, Xia Zeng, Lydia Dehbi, Zhengfeng Yang
摘要
Polynomial inequality proving is fundamental to many mathematical disciplines and finds wide applications in diverse fields. Current traditional algebraic methods are based on searching for a polynomial positive definite representation over a set of basis. However, these methods are limited by truncation degree. To address this issue, this paper proposes an approach based on reinforcement learning to find a Krivine-basis representation for proving polynomial inequalities. Specifically, we formulate the inequality proving problem as a linear programming (LP) problem and encode it as a basis selection problem using reinforcement learning (RL), achieving a non-negative Krivine basis. Moreover, a fast multivariate polynomial multiplication method based on Fast Fourier Transform (FFT) is employed to enhance the efficiency of action space search. Furthermore, we have implemented a tool called APPIRL (Automated Proof of Polynomial Inequalities via Reinforcement Learning). Experimental evaluation on benchmark problems demonstrates the feasibility and effectiveness of our approach. In addition, APPIRL has been successfully applied to solve the maximum stable set problem.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- RL-MUL: Multiplier Design Optimization with Deep Reinforcement LearningDongsheng Zuo, Yikang Ouyang, Yuzhe MaDAC 2023 · 被引用 17 次
- Learning Selection Strategies in Buchberger's AlgorithmDylan Peifer, Michael Eugene Stillman, Daniel Halpern-LeistnerICML 2020 · 被引用 35 次
- A Reinforcement-Learning-Based Multiple-Column Selection Strategy for Column GenerationHaofeng Yuan, Lichang Fang, Shiji SongAAAI 2024 · 被引用 11 次
- Accelerating Cutting-Plane Algorithms via Reinforcement Learning SurrogatesKyle Mana, Fernando Acero, Stephen Mak, Parisa Zehtabi 等AAAI 2024
- Reinforcement Learning for Integer Programming: Learning to CutYunhao Tang, Shipra Agrawal, Yuri FaenzaICML 2020 · 被引用 224 次
