Data-Efficient Learning with Neural Programs
Alaia Solko-Breslin, Seewon Choi, Ziyang Li, Neelay Velingker, Rajeev Alur, Mayur Naik, Eric Wong
Abstract
Many computational tasks can be naturally expressed as a composition of a DNN followed by a program written in a traditional programming language or an API call to an LLM. We call such composites"neural programs"and focus on the problem of learning the DNN parameters when the training data consist of end-to-end input-output labels for the composite. When the program is written in a differentiable logic programming language, techniques from neurosymbolic learning are applicable, but in general, the learning for neural programs requires estimating the gradients of black-box components. We present an algorithm for learning neural programs, called ISED, that only relies on input-output samples of black-box components. For evaluation, we introduce new benchmarks that involve calls to modern LLMs such as GPT-4 and also consider benchmarks from the neurosymbolic learning literature. Our evaluation shows that for the latter benchmarks, ISED has comparable performance to state-of-the-art neurosymbolic frameworks. For the former, we use adaptations of prior work on gradient approximations of black-box components as a baseline, and show that ISED achieves comparable accuracy but in a more data- and sample-efficient manner.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 24a70ae9-c7a8-4acb-b1d2-aeebf52fe2ffCited by top-tier papers4
- Lobster: A GPU-Accelerated Framework for Neurosymbolic ProgrammingPaul Biberstein, Ziyang Li, Joseph Devietti, Mayur NaikASPLOS 2026 · 1 citation
- CTSketch: Compositional Tensor Sketching for Scalable Neurosymbolic LearningSeewon Choi, Alaia Solko-Breslin, Rajeev Alur, Eric WongNeurIPS 2025 · 1 citation
- DOLPHIN: A Programmable Framework for Scalable Neurosymbolic LearningAaditya Naik, Jason Liu, Claire Wang, Amish Sethi et al.ICML 2025
- Once Upon an Input: Reasoning via Per-Instance Program SynthesisAdam Stein, Neelay Velingker, Mayur Naik, Eric WongNeurIPS 2025
Builds on9
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Scallop: From Probabilistic Deductive Databases to Scalable Differentiable ReasoningJiani Huang, Ziyang Li, Binghong Chen, Karan Samel et al.NeurIPS 2021 · 101 citations
- Closed Loop Neural-Symbolic Learning via Integrating Neural Perception, Grammar Parsing, and Symbolic ReasoningQing Li, Siyuan Huang, Yining Hong, Yixin Chen et al.ICML 2020 · 93 citations
- Neural-Symbolic Integration: A Compositional PerspectiveEfthymia Tsamoura, Timothy M. Hospedales, Loizos MichaelAAAI 2021 · 85 citations
- A-NeSI: A Scalable Approximate Method for Probabilistic Neurosymbolic InferenceEmile van Krieken, Thiviyan Thanapalasingam, Jakub M. Tomczak, Frank van Harmelen et al.NeurIPS 2023 · 62 citations
Related papers
- HYSYNTH: Context-Free LLM Approximation for Guiding Program SynthesisShraddha Barke, Emmanuel Anaya Gonzalez, Saketh Ram Kasibatla, Taylor Berg-Kirkpatrick et al.NeurIPS 2024 · 34 citations
- Large Language Models are Interpretable LearnersRuochen Wang, Si Si, Felix X. Yu, Dorothea Wiesmann Rothuizen et al.ICLR 2025
- From Perception to Programs: Regularize, Overparameterize, and AmortizeHao Tang, Kevin EllisICML 2023 · 13 citations
- Gradient-Based Program Synthesis with Neurally Interpreted LanguagesMatthew Macfarlane, Clément Bonnet, Herke van Hoof, Levi LelisICLR 2026 · 3 citations
- Web question answering with neurosymbolic program synthesisQiaochu Chen, Aaron Lamoreaux, Xinyu Wang, Greg Durrett et al.PLDI 2021 · 25 citations
