paper-with-me

홈 › Papers

Scaling Legal AI: Benchmarking Mamba and Transformers for Statutory Classification and Case Law Retrieval

2025-08-29 · Anuraj Maurya arxiv

The rapid growth of statutory corpora and judicial decisions requires scalable legal AI systems capable of classification and retrieval over extremely long contexts. Transformer-based architectures (e.g., Longformer, DeBERTa) dominate current legal NLP benchmarks but struggle with quadratic attention costs, limiting efficiency and scalability. In this work, we present the first comprehensive benchmarking of Mamba, a state-space model (SSM) with linear-time selective mechanisms, against leading transformer models for statutory classification and case law retrieval. We evaluate models on open-source legal corpora including LexGLUE, EUR-Lex, and ILDC, covering statutory tagging, judicial outcome prediction, and case retrieval tasks. Metrics include accuracy, recall at k, mean reciprocal rank (MRR), and normalized discounted cumulative gain (nDCG), alongside throughput measured in tokens per second and maximum context length. Results show that Mamba's linear scaling enables processing of legal documents several times longer than transformers, while maintaining or surpassing retrieval and classification performance. This study introduces a new legal NLP benchmark suite for long-context modeling, along with open-source code and datasets to support reproducibility. Our findings highlight trade-offs between state-space models and transformers, providing guidance for deploying scalable legal AI in statutory analysis, judicial decision support, and policy research.

📄 PDF Abstract BibTeX arXiv:2509.00141

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Bilingual BSARD: Extending Statutory Article Retrieval to Dutch

2024-12-10 · Ehsan Lotfi, Nikolay Banar, Nerses Yuzbashyan, Walter Daelemans

Statutory article retrieval plays a crucial role in making legal information more accessible to both laypeople and legal professionals. Multilingual countries like Belgium present unique challenges for retrieval models d…

ArticlesBenchmarkingRetrieval

Benchmarking Legal RAG: The Promise and Limits of AI Statutory Surveys

2026-02-07 · Mohamed Afane, Emaan Hariri, Derek Ouyang, Daniel E. Ho arxiv

Retrieval-augmented generation (RAG) offers significant potential for legal AI, yet systematic benchmarks are sparse. Prior work introduced LaborBench to benchmark RAG models based on ostensible ground truth from an exha…

LexKairos: Benchmarking Legal Temporal Capabilities in LLMs

2026-08-10 · Chenyang Li, Zejia Feng, Yuqin Huang, Yuxiao Ye 외 arxiv

Large language models (LLMs) have demonstrated strong performance across a wide range of legal tasks. In legal practice, time is a critical concept that governs the validity of statutes, the progression of legal cases, a…

Decompose-and-Refine: Structured Legal Question Answering with Parametric Retrieval

2026-05-23 · Jihyung lee, Hyounghun Kim, Gary Lee arxiv

Large language models (LLMs) have shown strong performance in the legal domain, demonstrating notable potential in Legal Question Answering (LQA). However, unlike general QA, LQA requires answers that are not only accura…

Question AnsweringLegal Reasoning

Bundesrecht: An Open Library and Corpus for German Statutory Reference Processing

2026-05-29 · Harshil Darji, Martin Heckelmann, Christina Kratsch, Gerard de Melo arxiv

Statutory references are central to legal language understanding, but are difficult to process automatically, as they appear in compact and variable surface forms, may combine multiple targets, use special abbreviations,…

Information Extraction