paper-with-me

홈 › Papers

SaulLM-7B: A pioneering Large Language Model for Law

2024-03-06 · Pierre Colombo, Telmo Pessoa Pires, Malik Boudiaf, Dominic Culver, Rui Melo, Caio Corro, Andre F. T. Martins, Fabrizio Esposito, Vera Lúcia Raposo, Sofia Morgado, Michael Desa

In this paper, we introduce SaulLM-7B, a large language model (LLM) tailored for the legal domain. With 7 billion parameters, SaulLM-7B is the first LLM designed explicitly for legal text comprehension and generation. Leveraging the Mistral 7B architecture as its foundation, SaulLM-7B is trained on an English legal corpus of over 30 billion tokens. SaulLM-7B exhibits state-of-the-art proficiency in understanding and processing legal documents. Additionally, we present a novel instructional fine-tuning method that leverages legal datasets to further enhance SaulLM-7B's performance in legal tasks. SaulLM-7B is released under the MIT License.

📄 PDF Abstract BibTeX arXiv:2403.03883

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingLarge Language ModelReading Comprehension

Similar Papers 제목 키워드 기반

SaulLM-54B & SaulLM-141B: Scaling Up Domain Adaptation for the Legal Domain

2024-07-28 · Pierre Colombo, Telmo Pires, Malik Boudiaf, Rui Melo 외

In this paper, we introduce SaulLM-54B and SaulLM-141B, two large language models (LLMs) tailored for the legal sector. These models, which feature architectures of 54 billion and 141 billion parameters, respectively, ar…

DecoderDomain AdaptationInstruction Following

Impacts of Continued Legal Pre-Training and IFT on LLMs' Latent Representations of Human-Defined Legal Concepts

2024-10-15 · Shaun Ho

This paper aims to offer AI & Law researchers and practitioners a more detailed understanding of whether and how continued pre-training and instruction fine-tuning (IFT) of large language models (LLMs) on legal corpora i…

The Factuality of Large Language Models in the Legal Domain

2024-09-18 · Rajaa El Hamdani, Thomas Bonald, Fragkiskos Malliaros, Nils Holzenberger 외

This paper investigates the factuality of large language models (LLMs) as knowledge bases in the legal domain, in a realistic usage scenario: we allow for acceptable variations in the answer, and let the model abstain fr…

Text to Trust: Evaluating Fine-Tuning and LoRA Trade-offs in Language Models for Unfair Terms of Service Detection

2025-10-26 · Noshitha Padma Pratyusha Juttu, Sahithi Singireddy, Sravani Gona, Sujal Timilsina arxiv

Large Language Models (LLMs) have transformed text understanding, yet their adaptation to specialized legal domains remains constrained by the cost of full fine-tuning. This study provides a systematic evaluation of fine…

DaG LLM ver 1.0: Pioneering Instruction-Tuned Language Modeling for Korean NLP

2023-11-23 · Dongjun Jang, Sangah Lee, Sungjoo Byun, Jinwoong Kim 외

This paper presents the DaG LLM (David and Goliath Large Language Model), a language model specialized for Korean and fine-tuned through Instruction Tuning across 41 tasks within 13 distinct categories.

Language ModelingLanguage ModellingLarge Language Model