Scallop: A Language for Neurosymbolic Programming
Ziyang Li, Jiani Huang, Mayur Naik
Abstract
We present Scallop, a language which combines the benefits of deep learning and logical reasoning. Scallop enables users to write a wide range of neurosymbolic applications and train them in a data- and compute-efficient manner. It achieves these goals through three key features: 1) a flexible symbolic representation that is based on the relational data model; 2) a declarative logic programming language that is based on Datalog and supports recursion, aggregation, and negation; and 3) a framework for automatic and efficient differentiable reasoning that is based on the theory of provenance semirings. We evaluate Scallop on a suite of eight neurosymbolic applications from the literature. Our evaluation demonstrates that Scallop is capable of expressing algorithmic reasoning in diverse and challenging AI tasks, provides a succinct interface for machine learning programmers to integrate logical domain knowledge, and yields solutions that are comparable or superior to state-of-the-art models in terms of accuracy. Furthermore, Scallop's solutions outperform these models in aspects such as runtime and data efficiency, interpretability, and generalizability.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ac646b98-d914-495e-8d85-23d33f2bf5f0Cited by top-tier papers25
- Exploiting Code Symmetries for Learning Program SemanticsKexin Pei, Weichen Li, Qirui Jin, Shuyang Liu et al.ICML 2024 · 15 citations
- Neurosymbolic Diffusion ModelsEmile van Krieken, Pasquale Minervini, Edoardo Maria Ponti, Antonio VergariNeurIPS 2025 · 12 citations
- Relational Programming with Foundational ModelsZiyang Li, Jiani Huang, Jason Liu, Felix Zhu et al.AAAI 2024 · 11 citations
- Data-Efficient Learning with Neural ProgramsAlaia Solko-Breslin, Seewon Choi, Ziyang Li, Neelay Velingker et al.NeurIPS 2024 · 10 citations
- ESCA: Contextualizing Embodied Agents via Scene-Graph GenerationJiani Huang, Amish Sethi, Matthew Kuo, Mayank Keoliya et al.NeurIPS 2025 · 7 citations
Builds on14
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 3,482 citations
- Long Range Arena : A Benchmark for Efficient TransformersYi Tay, Mostafa Dehghani, Samira Abnar, Yikang Shen et al.ICLR 2021 · 881 citations
- Neural Symbolic Reader: Scalable Integration of Distributed and Symbolic Representations for Reading ComprehensionXinyun Chen, Chen Liang, Adams Wei Yu, Denny Zhou et al.ICLR 2020 · 109 citations
- Learning Reasoning Strategies in End-to-End Differentiable ProvingPasquale Minervini, Sebastian Riedel, Pontus Stenetorp, Edward Grefenstette et al.ICML 2020 · 102 citations
Related papers
- Lobster: A GPU-Accelerated Framework for Neurosymbolic ProgrammingPaul Biberstein, Ziyang Li, Joseph Devietti, Mayur NaikASPLOS 2026 · 1 citation
- Scallop: From Probabilistic Deductive Databases to Scalable Differentiable ReasoningJiani Huang, Ziyang Li, Binghong Chen, Karan Samel et al.NeurIPS 2021 · 101 citations
- DOLPHIN: A Programmable Framework for Scalable Neurosymbolic LearningAaditya Naik, Jason Liu, Claire Wang, Amish Sethi et al.ICML 2025
- VAEL: Bridging Variational Autoencoders and Probabilistic Logic ProgrammingEleonora Misino, Giuseppe Marra, Emanuele SansoneNeurIPS 2022 · 38 citations
- DeepProofLog: Efficient Proving in Deep Stochastic Logic ProgramsYing Jiao, Rodrigo Castellano Ontiveros, Luc De Raedt, Marco Gori et al.AAAI 2026
