Factoring Statutory Reasoning as Language Understanding Challenges
Nils Holzenberger, Benjamin Van Durme
Abstract
Statutory reasoning is the task of determining whether a legal statute, stated in natural language, applies to the text description of a case. Prior work introduced a resource that approached statutory reasoning as a monolithic textual entailment problem, with neural baselines performing nearly at-chance. To address this challenge, we decompose statutory reasoning into four types of language-understanding challenge problems, through the introduction of concepts and structure found in Prolog programs. Augmenting an existing benchmark, we provide annotations for the four tasks, and baselines for three of them. Models for statutory reasoning are shown to benefit from the additional structure, improving on prior baselines. Further, the decomposition into subtasks facilitates finer-grained model diagnostics and clearer incremental progress.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d1cd2407-b0a2-4d3a-ba54-b49065ae81caCited by top-tier papers4
- Moving on from OntoNotes: Coreference Resolution Model TransferPatrick Xia, Benjamin Van DurmeEMNLP 2021 · 23 citations
- CourtReasoner: Can LLM Agents Reason Like Judges?Sophia Simeng Han, Yoshiki Takashima, Shannon Zejiang Shen, Chen Liu et al.EMNLP 2025 · 1 citation
- Lawma: The Power of Specialization for Legal AnnotationRicardo Dominguez-Olmedo, Vedant Nanda, Rediet Abebe, Stefan Bechtold et al.ICLR 2025 · 1 citation
- The Medium Is Not the Message: Deconfounding Document Embeddings via Linear Concept ErasureYu Fan, Yang Tian, Shauli Ravfogel, Mrinmaya Sachan et al.EMNLP 2025
Builds on3
- JEC-QA: A Legal-Domain Question Answering DatasetHaoxi Zhong, Chaojun Xiao, Cunchao Tu, Tianyang Zhang et al.AAAI 2020 · 212 citations
- Neural Module Networks for Reasoning over TextNitish Gupta, Kevin Lin, Dan Roth, Sameer Singh et al.ICLR 2020 · 134 citations
- Learning from Task DescriptionsOrion Weller, Nicholas Lourie, Matt Gardner, Matthew E. PetersEMNLP 2020 · 4 citations
Related papers
- LexGLUE: A Benchmark Dataset for Legal Language Understanding in EnglishIlias Chalkidis, Abhik Jana, Dirk Hartung, Michael J. Bommarito II et al.ACL 2022
- A Statutory Article Retrieval Dataset in FrenchAntoine Louis, Gerasimos SpanakisACL 2022 · 59 citations
- LexChain: Modeling Legal Reasoning Chains for Chinese Tort Case AnalysisHuiyuan Xie, Chenyang Li, Huining Zhu, Chubin Zhang et al.AAAI 2026 · 2 citations
- CLUES: A Benchmark for Learning Classifiers using Natural Language ExplanationsRakesh R. Menon, Sayan Ghosh, Shashank SrivastavaACL 2022 · 13 citations
- PLAWBENCH: A Rubric-Based Benchmark for Evaluating LLMs in Real-World Legal PracticeYuzhen Shi, Huanghai Liu, Yiran Hu, Gaojie Song et al.ACL 2026 · 7 citations
