AI-Assisted Moot Courts: Simulating Justice-Specific Questioning in Oral Arguments
In oral arguments, judges probe attorneys with questions about the factual record, legal claims, and the strength of their arguments. To prepare for this questioning, both law schools and practicing attorneys rely on moot courts: practice simulations of appellate hearings. Leveraging a dataset of U.S. Supreme Court oral argument transcripts, we examine whether AI models can effectively simulate justice-specific questioning for moot court-style training. Evaluating oral argument simulation is challenging because there is no single correct question for any given turn. Instead, effective questioning should reflect a combination of desirable qualities, such as anticipating substantive legal issues, detecting logical weaknesses, and maintaining an appropriately adversarial tone. We introduce a two-layer evaluation framework that assesses both the realism and pedagogical usefulness of simulated questions using complementary proxy metrics. We construct and evaluate both prompt-based and agentic oral argument simulators. We find that simulated questions are often perceived as realistic by human annotators and achieve high recall of ground truth substantive legal issues. However, models still face substantial shortcomings, including low diversity in question types and sycophancy. Importantly, these shortcomings would remain undetected under naive evaluation approaches.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
L'identification du locuteur : 20 ans de t\'emoignage dans les cours de Justice. Le cas du LIPSADON \textless\textless laboratoire ind\'ependant de police scientifique \textgreater\textgreater (Forensic speaker identification: 20 years of scientific testimonies in courts of Justice. The case of LIPSADON ``forensics independent laboratory'') [in French]
Racial Sentencing Disparities and Differential Progression Through the Criminal Justice System: Evidence From Linked Federal and State Court Data
Several key actors -- police, prosecutors, judges -- can alter the course of individuals passing through the multi-staged criminal justice system. I use linked arrest-sentencing data for federal courts from 1994-2010 to …
Selection biasBetter Transcription of UK Supreme Court Hearings
Transcription of legal proceedings is very important to enable access to justice. However, speech transcription is an expensive and slow process. In this paper we describe part of a combined research and industrial proje…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+2WCLD: Curated Large Dataset of Criminal Cases from Wisconsin Circuit Courts
Machine learning based decision-support tools in criminal justice systems are subjects of intense discussions and academic research. There are important open questions about the utility and fairness of such tools. Academ…
FairnessSentenceTathyaNyaya and FactLegalLlama: Advancing Factual Judgment Prediction and Explanation in the Indian Legal Context
In the landscape of Fact-based Judgment Prediction and Explanation (FJPE), reliance on factual data is essential for developing robust and realistic AI-driven decision-making tools. This paper introduces TathyaNyaya, the…
Decision MakingExplanation GenerationLarge Language Model