Learning Selection Strategies in Buchberger's Algorithm
Dylan Peifer, Michael Eugene Stillman, Daniel Halpern-Leistner
Abstract
Studying the set of exact solutions of a system of polynomial equations largely depends on a single iterative algorithm, known as Buchberger's algorithm. Optimized versions of this algorithm are crucial for many computer algebra systems (e.g., Mathematica, Maple, Sage). We introduce a new approach to Buchberger's algorithm that uses reinforcement learning agents to perform S-pair selection, a key step in the algorithm. We then study how the difficulty of the problem depends on the choices of domain and distribution of polynomials, about which little is known. Finally, we train a policy model using proximal policy optimization (PPO) to learn S-pair selection strategies for random systems of binomial equations. In certain domains, the trained model outperforms state-of-the-art selection heuristics in total number of polynomial additions performed, which provides a proof-of-concept that recent developments in machine learning have the potential to improve performance of algorithms in symbolic computation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 82aef21f-70e7-4e91-b5d7-a99979b58c40Cited by top-tier papers4
- Learning to compute Gröbner basesHiroshi Kera, Yuki Ishihara, Yuta Kambe, Tristan Vaccon et al.NeurIPS 2024 · 9 citations
- Computational Algebra with Attention: Transformer Oracles for Border Basis AlgorithmsHiroshi Kera, Nico Pelleriti, Yuki Ishihara, Max Zimmer et al.NeurIPS 2025 · 8 citations
- Applying language models to algebraic topology: generating simplicial cycles using multi-labeling in Wu's formulaKirill Brilliantov, Fedor Pavutnitskiy, Dmitry Pasechnyuk, German MagaiICML 2024 · 1 citation
- HATSolver: Learning Gröbner Bases with Hierarchical Attention TransformersMohamed Malhou, Ludovic Perret, Kristin E. LauterICLR 2026
Builds on1
Related papers
- Learning Branching Policies for MILPs with Proximal Policy OptimizationAbdelouahed Ben Mhamed, Assia Kamal Idrissi, Amal El Fallah SeghrouchniAAAI 2026
- Automated Proof of Polynomial Inequalities via Reinforcement LearningBanglong Liu, Niuniu Qi, Xia Zeng, Lydia Dehbi et al.CVPR 2025
- Model-based reinforcement learning for biological sequence designChristof Angermüller, David Dohan, David Belanger, Ramya Deshpande et al.ICLR 2020 · 159 citations
- Planning in Branch-and-Bound: Model-Based Reinforcement Learning for Exact Combinatorial OptimizationPaul Strang, Zacharie Alès, Côme Bissuel, Olivier Juan et al.AAAI 2026
- Improving Exact Algorithm for Pseudo Boolean Optimization with Two New Phase Selection HeuristicsYujiao Zhao, Yizhan Xiang, Jiangnan Li, Yiyuan Wang et al.AAAI 2026
