LeNER-Br: a Dataset for Named Entity Recognition in Brazilian Legal Text
Named entity recognition systems have the untapped potential to extract information from legal documents, which can improve information retrieval and decision-making processes. In this paper, a dataset for named entity recognition in Brazilian legal documents is presented. Unlike other Portuguese language datasets, this dataset is composed entirely of legal documents. In addition to tags for persons, locations, time entities and organizations, the dataset contains specific tags for law and legal cases entities. To establish a set of baseline results, we first performed experiments on another Portuguese dataset: Paramopama. This evaluation demonstrate that LSTM-CRF gives results that are significantly better than those previously reported. We then retrained LSTM-CRF, on our dataset and obtained F 1 scores of 97.04% and 88.82% for Legislation and Legal case entities, respectively. These results show the viability of the proposed dataset for legal applications.
Code (1)
Tasks
Decision MakingInformation Retrievalnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)RetrievalSimilar Papers 제목 키워드 기반
Cross-Lingual Named Entity Recognition via FastAlign: a Case Study
Named Entity Recognition is an essential task in natural language processing to detect entities and classify them into predetermined categories. An entity is a meaningful word, or phrase that refers to proper nouns. Name…
Machine Translationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+4UlyssesNER-Br: A Corpus of Brazilian Legislative Documents for Named Entity Recognition
The amount of legislative documents produced within the past decade has risen dramatically, making it difficult for law practitioners to consult and update legislation. Named Entity Recognition (NER) systems have the unt…
Decision MakingInformation Retrievalnamed-entity-recognitionNamed Entity Recognition+3Evaluating Named Entity Recognition: A comparative analysis of mono- and multilingual transformer models on a novel Brazilian corporate earnings call transcripts dataset
Since 2018, when the Transformer architecture was introduced, Natural Language Processing has gained significant momentum with pre-trained Transformer-based models that can be fine-tuned for various tasks. Most models ar…
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+3CDJUR-BR -- A Golden Collection of Legal Document from Brazilian Justice with Fine-Grained Named Entities
A basic task for most Legal Artificial Intelligence (Legal AI) applications is Named Entity Recognition (NER). However, texts produced in the context of legal practice make references to entities that are not trivially r…
AttributeJurisprudencenamed-entity-recognitionNamed Entity Recognition+2Artificial Intelligence in Brazilian News: A Mixed-Methods Analysis
The current surge in Artificial Intelligence (AI) interest, reflected in heightened media coverage since 2009, has sparked significant debate on AI's implications for privacy, social justice, workers' rights, and democra…
Articlesnamed-entity-recognitionNamed Entity Recognition