On Verifiable Legal Reasoning: A Multi-Agent Framework with Formalized Knowledge Representations
arXiv:2509.00710 · doi:10.1145/3746252.3761057
Abstract
Legal reasoning requires both precise interpretation of statutory language and consistent application of complex rules, presenting significant challenges for AI systems. This paper introduces a modular multi-agent framework that decomposes legal reasoning into distinct knowledge acquisition and application stages. In the first stage, specialized agents extract legal concepts and formalize rules to create verifiable intermediate representations of statutes. The second stage applies this knowledge to specific cases through three steps: analyzing queries to map case facts onto the ontology schema, performing symbolic inference to derive logically entailed conclusions, and generating final answers using a programmatic implementation that operationalizes the ontological knowledge. This bridging of natural language understanding with symbolic reasoning provides explicit and verifiable inspection points, significantly enhancing transparency compared to end-to-end approaches. Evaluation on statutory tax calculation tasks demonstrates substantial improvements, with foundational models achieving 76.4\% accuracy compared to 18.8\% baseline performance, effectively narrowing the performance gap between reasoning and foundational models. These findings suggest that modular architectures with formalized knowledge representations can make sophisticated legal reasoning more accessible through computationally efficient models while enhancing consistency and explainability in AI legal reasoning, establishing a foundation for future research into more transparent, trustworthy, and effective AI systems for legal domain.
Accepted for publication at the 34th ACM International Conference on Information and Knowledge Management (CIKM '25)
References in corpus (14)
- Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
- LegalBench: Prototyping a Collaborative Benchmark for Legal Reasoning
- LLMs4OL: Large Language Models for Ontology Learning
- Towards Ontology Construction with Language Models
- Thought-Like-Pro: Enhancing Reasoning of Large Language Models through Self-Bootstrapped Prolog-based Chain-of-Thought
- Autoformalizing Natural Language to First-Order Logic: A Case Study in Logical Fallacy Detection
- Chain of Logic: Rule-Based Reasoning with Large Language Models
- Can LLMs Follow Simple Rules?
- Leverage Knowledge Graph and Large Language Model for Law Article Recommendation: A Case Study of Chinese Criminal Law
- LLMs Provide Unstable Answers to Legal Questions
- Logically Consistent Language Models via Neuro-Symbolic Integration
- Can Large Language Models Grasp Legal Theories? Enhance Legal Reasoning with Insights from Multi-Agent Collaboration
- Teaching AI to Handle Exceptions: Supervised Fine-Tuning with Human-Aligned Judgment
- Towards Robust Legal Reasoning: Harnessing Logical LLMs in Law