From Dissonance to Insights: Dissecting Disagreements in Rationale Construction for Case Outcome Classification
Shanshan Xu, T. Y. S. S. Santosh, Oana Ichim, Isabella Risini, Barbara Plank, Matthias Grabmair
摘要
In legal NLP, Case Outcome Classification<br/>(COC) must not only be accurate but also<br/>trustworthy and explainable. Existing work<br/>in explainable COC has been limited to an-<br/>notations by a single expert. However, it is<br/>well-known that lawyers may disagree in their<br/>assessment of case facts. We hence collect<br/>a novel dataset RAVE: Rationale Variation<br/>in ECHR1, which is obtained from two ex-<br/>perts in the domain of international human<br/>rights law, for whom we observe weak agree-<br/>ment. We study their disagreements and build a<br/>two-level task-independent taxonomy, supple-<br/>mented with COC-specific subcategories. We<br/>quantitatively assess different taxonomy cate-<br/>gories and find that disagreements mainly stem<br/>from underspecification of the legal context,<br/>which poses challenges given the typically lim-<br/>ited granularity and noise in COC metadata. To<br/>our knowledge, this is the first work in the legal<br/>NLP that focuses on building a taxonomy over<br/>human label variation. We further assess the ex-<br/>plainablility of state-of-the-art COC models on<br/>RAVE and observe limited agreement between<br/>models and experts. Overall, our case study re-<br/>veals hitherto underappreciated complexities in<br/>creating benchmark datasets in legal NLP that<br/>revolve around identifying aspects of a case’s<br/>facts supposedly relevant to its outcome
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Which Demographics do LLMs Default to During Annotation?Johannes Schäfer, Aidan Combs, Christopher Bagdon, Jiahui Li 等ACL 2025 · 被引用 11 次
- Through the Lens of Split Vote: Exploring Disagreement, Difficulty and Calibration in Legal Case Outcome ClassificationShanshan Xu, T. Y. S. S. Santosh, Oana Ichim, Barbara Plank 等ACL 2024
它引用的顶会 Paper4
- NeurJudge: A Circumstance-aware Neural Framework for Legal Judgment PredictionLinan Yue, Qi Liu, Binbin Jin, Han Wu 等SIGIR 2021 · 被引用 83 次
- Deconfounding Legal Judgment Prediction for European Court of Human Rights Cases Towards Better Alignment with ExpertsTokala Yaswanth Sri Sai Santosh, Shanshan Xu, Oana Ichim, Matthias GrabmairEMNLP 2022 · 被引用 13 次
- ILDC for CJPE: Indian Legal Documents Corpus for Court Judgment Prediction and ExplanationVijit Malik, Rishabh Sanjay, Shubham Kumar Nigam, Kripabandhu Ghosh 等ACL 2021
- LexGLUE: A Benchmark Dataset for Legal Language Understanding in EnglishIlias Chalkidis, Abhik Jana, Dirk Hartung, Michael J. Bommarito II 等ACL 2022
相关 Paper
- Explainable Legal Case Matching via Inverse Optimal Transport-based Rationale ExtractionWeijie Yu, Zhongxiang Sun, Jun Xu, Zhenhua Dong 等SIGIR 2022 · 被引用 45 次
- Evaluating Legal Reasoning Traces with Legal Issue Tree RubricsJinu Lee, Kyoung-Woon On, Sophia Simeng Han, Arman Cohan 等ACL 2026 · 被引用 2 次
- CasiMedicos-Arg: A Medical Question Answering Dataset Annotated with Explanatory Argumentative StructuresEkaterina Sviridova, Anar Yeginbergen, Ainara Estarrona, Elena Cabrio 等EMNLP 2024 · 被引用 2 次
- e-CARE: a New Dataset for Exploring Explainable Causal ReasoningLi Du, Xiao Ding, Kai Xiong, Ting Liu 等ACL 2022
- Legal Fact Prediction: The Missing Piece in Legal Judgment PredictionJunkai Liu, Yujie Tong, Hui Huang, Bowen Zheng 等EMNLP 2025
