paper-with-me

홈 › Papers

InsufficiencyBench: Evaluating LLM legal advice on underspecified user queries

2026-08-20 · Samuel J. Vincent, Daniel Calloway, Fangyi Yu, Andrew M. Bean, Nabeel Seedat arxiv

Legal AI systems are increasingly used to answer legal questions, yet existing benchmarks assume queries arrive fully specified. In practice, users omit facts that materially determine the legal outcome. We introduce InsufficiencyBench, the first legal benchmark targeting query-side insufficiency: whether a model recognizes when a query lacks legally material information, identifies what is missing, and refrains from premature conclusions. We formalize a taxonomy of eight canonical missing-element categories across three structural failure modes---switch, gating, and fatal prerequisite--- and construct 202 benchmark items (58 base queries, 144 deficient variants) spanning six legal domains and 24 US jurisdictions and annotated by practising attorneys. Evaluating ten frontier models, we find that no model exceeds F2 = 0.46 on missing-element identification and that the median recall is 0.44. Models either hedge indiscriminately or answer silently under fabricated presumptions. No model both identifies and qualifies responses to deficient queries while directly addressing complete ones.

📄 PDF Abstract BibTeX arXiv:2608.20220

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Which Institutional Frameworks Do Chatbots Assume? Auditing Jurisdictional Defaults in Multilingual LLMs

2026-05-29 · Zhizhi Wang, Harini Suresh arxiv

LLMs increasingly answer questions about taxes, labor protections, healthcare, education, pensions, and administrative procedures, where usefulness often depends on the applicable jurisdiction. Multilingual users may wri…

(A)I Am Not a Lawyer, But...: Engaging Legal Experts towards Responsible LLM Policies for Legal Advice

2024-02-02 · Inyoung Cheong, King Xia, K. J. Kevin Feng, Quan Ze Chen 외

Large language models (LLMs) are increasingly capable of providing users with advice in a wide range of professional domains, including legal advice. However, relying on LLMs for legal queries raises concerns due to the …

Ethics

ELLA: Empowering LLMs for Interpretable, Accurate and Informative Legal Advice

2024-08-13 · Yutong Hu, Kangcheng Luo, Yansong Feng

Despite remarkable performance in legal consultation exhibited by legal Large Language Models(LLMs) combined with legal article retrieval components, there are still cases when the advice given is incorrect or baseless. …

Articles

Answer Retrieval in Legal Community Question Answering

2024-01-09 · Arian Askari, Zihui Yang, Zhaochun Ren, Suzan Verberne

The task of answer retrieval in the legal domain aims to help users to seek relevant legal advice from massive amounts of professional responses. Two main challenges hinder applying existing answer retrieval approaches i…

Community Question AnsweringQuestion AnsweringRetrieval

A Brief Report on LawGPT 1.0: A Virtual Legal Assistant Based on GPT-3

2023-02-11 · Ha-Thanh Nguyen

LawGPT 1.0 is a virtual legal assistant built on the state-of-the-art language model GPT-3, fine-tuned for the legal domain. The system is designed to provide legal assistance to users in a conversational manner, helping…

Language ModelingLanguage Modelling