paper-with-me

Papers

Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models

2024-01-02 · Matthew Dahl, Varun Magesh, Mirac Suzgun, Daniel E. Ho

Do large language models (LLMs) know the law? These models are increasingly being used to augment legal practice, education, and research, yet their revolutionary potential is threatened by the presence of hallucinations -- textual output that is not consistent with legal facts. We present the first systematic evidence of these hallucinations, documenting LLMs' varying performance across jurisdictions, courts, time periods, and cases. Our work makes four key contributions. First, we develop a typology of legal hallucinations, providing a conceptual framework for future research in this area. Second, we find that legal hallucinations are alarmingly prevalent, occurring between 58% of the time with ChatGPT 4 and 88% with Llama 2, when these models are asked specific, verifiable questions about random federal court cases. Third, we illustrate that LLMs often fail to correct a user's incorrect legal assumptions in a contra-factual question setup. Fourth, we provide evidence that LLMs cannot always predict, or do not always know, when they are producing legal hallucinations. Taken together, our findings caution against the rapid and unsupervised integration of popular LLMs into legal tasks. Even experienced lawyers must remain wary of legal hallucinations, and the risks are highest for those who stand to benefit from LLMs the most -- pro se litigants or those without access to traditional legal resources.

📄 PDF Abstract BibTeX arXiv:2401.01301

Code (1)

reglab/legal_hallucinations 공식 구현

Similar Papers 제목 키워드 기반

LexAgentHallu: A Hierarchical Benchmark for Profiling Hallucinations in Legal Agents

2026-09-09 · Yujin Zhou, Mingxuan Zheng, Chuxue Cao, Huang Yidan 외 arxiv

As large language models are increasingly deployed as tool-augmented legal agents, they introduce agentic hallucinations where tool-call and reasoning errors cascade into fabricated holdings and miscited authority. Howev…

Gaps or Hallucinations? Gazing into Machine-Generated Legal Analysis for Fine-grained Text Evaluations

2024-09-16 · Abe Bohan Hou, William Jurayj, Nils Holzenberger, Andrew Blair-Stanek 외

Large Language Models (LLMs) show promise as a writing aid for professionals performing legal analyses. However, LLMs can often hallucinate in this setting, in ways difficult to recognize by non-professionals and existin…

Chatlaw: A Multi-Agent Collaborative Legal Assistant with Knowledge Graph Enhanced Mixture-of-Experts Large Language Model

2023-06-28 · Jiaxi Cui, Munan Ning, Zongjian Li, Bohua Chen 외

AI legal assistants based on Large Language Models (LLMs) can provide accessible legal consulting services, but the hallucination problem poses potential legal risks. This paper presents Chatlaw, an innovative legal assi…

HallucinationKnowledge GraphsLanguage ModelingLanguage Modelling+3

Hallucination-Free? Assessing the Reliability of Leading AI Legal Research Tools

2024-05-30 · Varun Magesh, Faiz Surani, Matthew Dahl, Mirac Suzgun 외

Legal practice has witnessed a sharp rise in products incorporating artificial intelligence (AI). Such tools are designed to assist with a wide range of core legal tasks, from search and summarization of caselaw to docum…

HallucinationRAGRetrieval-augmented Generation

Artificial Intelligence and Legal Analysis: Implications for Legal Education and the Profession

2025-02-04 · Lee Peoples

This article reports the results of a study examining the ability of legal and non-legal Large Language Models to perform legal analysis using the Issue-Rule-Application-Conclusion framework. LLMs were tested on legal re…

Legal Reasoning