Learning Interestingness in Automated Mathematical Theory Formation
George Tsoukalas, Rahul Saha, Amitayush Thakur, Sabrina Reguyal, Swarat Chaudhuri
摘要
We take two key steps in automating the open-ended discovery of new mathematical theories, a grand challenge in artificial intelligence. First, we introduce , a reinforcement learning (RL) environment that models concept discovery and theorem-proving using a set of symbolic actions, opening up a range of RL problems relevant to theory discovery. Second, we explore a specific problem through : automatically scoring the of mathematical objects. We investigate evolutionary algorithms for synthesizing nontrivial interestingness measures. In particular, we introduce an LLM-based evolutionary algorithm that features function abstraction, leading to notable improvements in discovering elementary number theory and finite fields over hard-coded baselines. We open-source the environment at this URL(https://github.com/trishullab/Fermat).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- Eureka: Human-Level Reward Design via Coding Large Language ModelsYecheng Jason Ma, William Liang, Guanzhi Wang, De-An Huang 等ICLR 2024 · 被引用 582 次
- Symbolic Regression with a Learned Concept LibraryArya Grayeli, Atharva Sehgal, Omar Costilla-Reyes, Miles D. Cranmer 等NeurIPS 2024 · 被引用 105 次
- Meta-learning curiosity algorithmsFerran Alet, Martin F. Schneider, Tomás Lozano-Pérez, Leslie Pack KaelblingICLR 2020 · 被引用 67 次
- Learning Formal Mathematics From Intrinsic MotivationGabriel Poesia, David Broman, Nick Haber, Noah D. GoodmanNeurIPS 2024 · 被引用 55 次
- Top-Down Synthesis for Library LearningMatthew Bowers, Theo X. Olausson, Lionel Wong, Gabriel Grand 等POPL 2023 · 被引用 32 次
相关 Paper
- DecAEvolve: Decompose, Adapt, and Evolve for Effective LLM-based Scientific Equation DiscoveryPouya Behzadifar, Parshin Shojaee, Sanchit Kabra, Kazem Meidani 等ICML 2026
- A Deep Reinforcement Learning Approach to First-Order Logic Theorem ProvingMaxwell Crouse, Ibrahim Abdelaziz, Bassem Makni, Spencer Whitehead 等AAAI 2021 · 被引用 41 次
- TacticZero: Learning to Prove Theorems from Scratch with Deep Reinforcement LearningMinchao Wu, Michael Norrish, Christian Walder, Amir DezfouliNeurIPS 2021 · 被引用 56 次
- Learning to Prove Theorems by Learning to Generate TheoremsMingzhe Wang, Jia DengNeurIPS 2020 · 被引用 60 次
- Agentic RL Scaling Law: Spontaneous Code Execution for Mathematical Problem SolvingXinji Mai, Haotian Xu, Xing W, Weinong Wang 等NeurIPS 2025 · 被引用 7 次
