paper-with-me

Papers

ReGal: A First Look at PPO-based Legal AI for Judgment Prediction and Summarization in India

2025-12-19 · Shubham Kumar Nigam, Tanuj Tyagi, Siddharth Shukla, Aditya Kumar Guru, Balaramamahanthi Deepak Patnaik, Danush Khanna, Noel Shallum, Kripabandhu Ghosh, Arnab Bhattacharya arxiv

This paper presents an early exploration of reinforcement learning methodologies for legal AI in the Indian context. We introduce Reinforcement Learning-based Legal Reasoning (ReGal), a framework that integrates Multi-Task Instruction Tuning with Reinforcement Learning from AI Feedback (RLAIF) using Proximal Policy Optimization (PPO). Our approach is evaluated across two critical legal tasks: (i) Court Judgment Prediction and Explanation (CJPE), and (ii) Legal Document Summarization. Although the framework underperforms on standard evaluation metrics compared to supervised and proprietary models, it provides valuable insights into the challenges of applying RL to legal texts. These challenges include reward model alignment, legal language complexity, and domain-specific adaptation. Through empirical and qualitative analysis, we demonstrate how RL can be repurposed for high-stakes, long-document tasks in law. Our findings establish a foundation for future work on optimizing legal reasoning pipelines using reinforcement learning, with broader implications for building interpretable and adaptive legal AI systems.

📄 PDF Abstract BibTeX arXiv:2512.18014

Code (0)

등록된 구현이 없습니다.

Tasks

Document SummarizationReinforcement LearningLegal Reasoning

Similar Papers 제목 키워드 기반

LegalDuet: Learning Fine-grained Representations for Legal Judgment Prediction via a Dual-View Contrastive Learning

2024-01-27 · Buqiang Xu, Xin Dai, Zhenghao Liu, Huiyuan Xie 외

Legal Judgment Prediction (LJP) is a fundamental task of legal artificial intelligence, aiming to automatically predict the judgment outcomes of legal cases. Existing LJP models primarily focus on identifying legal trigg…

Contrastive Learning

Legal Fact Prediction: The Missing Piece in Legal Judgment Prediction

2024-09-11 · Junkai Liu, Yujie Tong, Hui Huang, Bowen Zheng 외

Legal judgment prediction (LJP), which enables litigants and their lawyers to forecast judgment outcomes and refine litigation strategies, has emerged as a crucial legal NLP task. Existing studies typically utilize legal…

Prediction

CAIL2018: A Large-Scale Legal Dataset for Judgment Prediction

2018-07-04 · Chaojun Xiao, Haoxi Zhong, Zhipeng Guo, Cunchao Tu 외

In this paper, we introduce the \textbf{C}hinese \textbf{AI} and \textbf{L}aw challenge dataset (CAIL2018), the first large-scale Chinese legal dataset for judgment prediction. \dataset contains more than $2.6$ million c…

ArticlesPredictionText Classification

AppealCase: A Dataset and Benchmark for Civil Case Appeal Scenarios

2025-05-22 · YuTing Huang, Meitong Guo, Yiquan Wu, Ang Li 외

Recent advances in LegalAI have primarily focused on individual case judgment analysis, often overlooking the critical appellate process within the judicial system. Appeals serve as a core mechanism for error correction …

Decision MakingMulti-class ClassificationText ClassificationText Generation

Explicitly Integrating Judgment Prediction with Legal Document Retrieval: A Law-Guided Generative Approach

2023-12-15 · Weicong Qin, Zelin Cao, Weijie Yu, Zihua Si 외

Legal document retrieval and judgment prediction are crucial tasks in intelligent legal systems. In practice, determining whether two documents share the same judgments is essential for establishing their relevance in le…

PredictionRetrievalSemantic SimilaritySemantic Textual Similarity