PAKTON: A Multi-Agent Framework for Question Answering in Long Legal Agreements
Petros Raptopoulos, Giorgos Filandrianos, Maria Lymperaiou, Giorgos Stamou
摘要
Contract review is a complex and timeintensive task that typically demands specialized legal expertise, rendering it largely inaccessible to non-experts. Moreover, legal interpretation is rarely straightforward-ambiguity is pervasive, and judgments often hinge on subjective assessments. Compounding these challenges, contracts are usually confidential, restricting their use with proprietary models and necessitating reliance on open-source alternatives. To address these challenges, we introduce PAKTON: a fully open-source, endto-end, multi-agent framework with plug-andplay capabilities. PAKTON is designed to handle the complexities of contract analysis through collaborative agent workflows and a novel multi-stage retrieval-augmented generation (RAG) component, enabling automated legal document review that is more accessible, adaptable, and privacy-preserving. Experiments demonstrate that PAKTON outperforms both general-purpose and pretrained models in predictive accuracy, retrieval performance, explainability, completeness, and grounded justifications as evaluated through a human study and validated with automated metrics. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper12
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni 等NeurIPS 2020 · 被引用 19,162 次
- G-Eval: NLG Evaluation using Gpt-4 with Better Human AlignmentYang Liu, Dan Iter, Yichong Xu, Shuohang Wang 等EMNLP 2023 · 被引用 549 次
- Can Large Language Models Be an Alternative to Human Evaluations?David Cheng-Han Chiang, Hung-yi LeeACL 2023 · 被引用 254 次
相关 Paper
- ProvBench: A Benchmark of Legal Provision Recommendation for Contract Auto-ReviewingXiuxuan Shen, Zhongyuan Jiang, Junsan Zhang, Junxiao Han 等ACL 2025 · 被引用 2 次
- ACORD: An Expert-Annotated Retrieval Dataset for Legal Contract DraftingSteven H. Wang, Maksim Zubkov, Kexin Fan, Sarah Harrell 等ACL 2025 · 被引用 14 次
- LegalGraphRAG: Multi-Agent Graph Retrieval-Augmented Generation for Reliable Legal ReasoningZerui Chen, Qinggang Zhang, Zhishang Xiang, Zhimin Wei 等ACL 2026 · 被引用 2 次
- DocETL: Agentic Query Rewriting and Evaluation for Complex Document ProcessingShreya Shankar, Tristan Chambers, Tarak Shah, Aditya G. Parameswaran 等VLDB 2025 · 被引用 62 次
- Automating Legal Interpretation with LLMs: Retrieval, Generation, and EvaluationKangcheng Luo, Quzhe Huang, Cong Jiang, Yansong FengACL 2025 · 被引用 4 次
