JEC-QA: A Legal-Domain Question Answering Dataset
Haoxi Zhong, Chaojun Xiao, Cunchao Tu, Tianyang Zhang, Zhiyuan Liu, Maosong Sun
Abstract
We present JEC-QA, the largest question answering dataset in the legal domain, collected from the National Judicial Examination of China. The examination is a comprehensive evaluation of professional skills for legal practitioners. College students are required to pass the examination to be certified as a lawyer or a judge. The dataset is challenging for existing question answering methods, because both retrieving relevant materials and answering questions require the ability of logic reasoning. Due to the high demand of multiple reasoning abilities to answer legal questions, the state-of-the-art models can only achieve about 28% accuracy on JEC-QA, while skilled humans and unskilled humans can reach 81% and 64% accuracy respectively, which indicates a huge gap between humans and machines on this task. We will release JEC-QA and our baselines to help improve the reasoning ability of machine comprehension models. You can access the dataset from http://jecqa.thunlp.org/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6bb0e5a5-dcdd-47e3-b78e-eeaa4d4b82ddCited by top-tier papers42
- How Does NLP Benefit Legal System: A Summary of Legal Artificial IntelligenceHaoxi Zhong, Chaojun Xiao, Cunchao Tu, Tianyang Zhang et al.ACL 2020 · 316 citations
- At Which Training Stage Does Code Data Help LLMs Reasoning?Yingwei Ma, Yue Liu, Yue Yu, Yuanliang Zhang et al.ICLR 2024 · 106 citations
- Judgment Prediction via Injecting Legal Knowledge into Neural NetworksLeilei Gan, Kun Kuang, Yi Yang, Fei WuAAAI 2021 · 72 citations
- A Statutory Article Retrieval Dataset in FrenchAntoine Louis, Gerasimos SpanakisACL 2022 · 59 citations
- LawBench: Benchmarking Legal Knowledge of Large Language ModelsZhiwei Fei, Xiaoyu Shen, Dawei Zhu, Fengzhe Zhou et al.EMNLP 2024 · 59 citations
Related papers
- MLEC-QA: A Chinese Multi-Choice Biomedical Question Answering DatasetJing Li, Shangping Zhong, Kaizhi ChenEMNLP 2021 · 24 citations
- ReCO: A Large Scale Chinese Reading Comprehension Dataset on OpinionBingning Wang, Ting Yao, Qi Zhang, Jingfang Xu et al.AAAI 2020 · 26 citations
- From Query to Counsel: Structured Reasoning with a Multi-Agent Framework and Dataset for Legal ConsultationMingfei Lu, Yi Zhang, Mengjia Wu, Yue FengACL 2026 · 2 citations
- LEXam: Benchmarking Legal Reasoning on 340 Law ExamsYu Fan, Jingwei Ni, Jakob Merane, Yang Tian et al.ICLR 2026 · 56 citations
- Towards Medical Machine Reading Comprehension with Structural Knowledge and Plain TextDongfang Li, Baotian Hu, Qingcai Chen, Weihua Peng et al.EMNLP 2020 · 39 citations
