paper-with-me

홈 › Papers

Can Legal AI Know When It Is Wrong? And Do Students Know When It Is?

2026-08-21 · Angel Mary John, Vipin Kumar Singh, Jerrin Thomas Panachakel arxiv

Integrating Large Language Models (LLMs) into the Indian judiciary promises access to justice but introduces severe risks. We identify the 'inertia of confidence'--an overconfidence phenomenon analogous to the Dunning-Kruger effect where LLMs provide incorrect legal verdicts with near-maximum confidence, driven by a hypothesized 'precedent overfitting' bias. Phase I of our socio-technical audit tested ChatGPT (GPT-5.2), Meta AI, and Perplexity AI on a 60-case battery regarding the Indian Contract Act, 1872, and the shift toward statutory enforcement of specific performance. We introduce the High-Confidence Error Rate (HCER) to quantify incorrect verdicts delivered with dangerous certainty (>= 9 on a 1-10 scale). All models struggled with statutory updates. Meta AI proved most vulnerable (31.7% HCER), frequently misapplying pre-amendment rules with a 9.1/10 mean confidence, followed by Perplexity (15.0%) and ChatGPT (6.7%). Phase II investigated human vulnerability to this overconfidence via a survey of Indian law students (N=380). Verification often functions as a reactive adaptation to machine hallucinations: students encountering fabricated citations reported higher verification scores (4.2/5) than those with no such encounters (2.8/5). Furthermore, while 81.6% knew submitting hallucinated cases can lead to contempt-of-court, 71.1% received no formal training on ethical AI use. We propose shifting toward adversarial legal research pedagogy and implementing source-grounded verification architectures to prevent systemic professional negligence.

📄 PDF Abstract BibTeX arXiv:2608.21089

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CausalOPD: First-Wrong-Step Supervision for Distilling Causal Chain Reasoning

2026-08-04 · Jian Zhang, Bingyi Wang, Yizhi Liu arxiv

Many critical reasoning tasks, including clinical diagnosis, legal judgment, and industrial fault diagnosis, require step-dependent causal chains in which early errors propagate and correct conclusions can mask invalid r…

Reinforcement LearningFault Diagnosis

Eskwai for Students: Generative AI Assistant for Legal Education in Ghana

2026-05-14 · George Boateng, Philemon Badu, Patrick Agyeman-Budu, Samuel Ansah 외 arxiv

Recent advances in generative AI have shown their potential to be leveraged for legal education. Yet, work on the development and deployment of such systems for legal education in the Global South is limited. In this wor…

Is this Citation on Point?

2026-08-12 · Apurv Verma hf

In 2023, a New York judge sanctioned two attorneys in Mata v. Avianca for filing a brief with hallucinated citations generated by ChatGPT. Such failures are largely caught by database lookups; the harder problem is detec…

Do Not Trust a Model Because It is Confident: Uncovering and Characterizing Unknown Unknowns to Student Success Predictors in Online-Based Learning

2022-12-16 · Roberta Galici, Tanja Käser, Gianni Fenu, Mirko Marras

Student success models might be prone to develop weak spots, i.e., examples hard to accurately classify due to insufficient representation during model creation. This weakness is one of the main factors undermining users…

Informativeness

Asymmetric Temperature Scaling Makes Larger Networks Teach Well Again

2022-10-10 · Xin-Chun Li, Wen-Shu Fan, Shaoming Song, Yinchuan Li 외

Knowledge Distillation (KD) aims at transferring the knowledge of a well-performed neural network (the {\it teacher}) to a weaker one (the {\it student}). A peculiar phenomenon is that a more accurate model doesn't neces…

Knowledge Distillation