paper-with-me

홈 › Papers

Legal Terminology Extraction with the Termolator

2021-11-01 · EMNLP (NLLP) 2021 11 · Nhi Pham, Lachlan Pham, Adam L. Meyers

Domain-specific terminology is ubiquitous in legal documents. Despite potential utility in populating glossaries and ontologies or as arguments in information extraction and document classification tasks, there has been limited work done for legal terminology extraction. This paper describes some work to remedy this omission. In the described research, we make some modifications to the Termolator, a high-performing, open-source terminology extractor which has been tuned to scientific articles. Our changes are designed to improve the Termolator’s results when applied to United States Supreme Court decisions. Unaltered and using the recommended settings, the original Termolator provides a list of terminology with a precision of 23% and 25% for the categories of economic activity (development set) and criminal procedures (test set) respectively. These were the most frequently occurring broad issues in Washington University in St. Louis Database corpus, a database of Supreme Court decisions that have been manually classified by topic. Our contribution includes the introduction of several legal domain-specific filtration steps and changes to the web search relevance score; each incrementally improved precision culminating in a combined precision of 63% and 65%. We also evaluated the baseline version of the Termolator on more specific subcategories and on broad issues with fewer cases. Our results show that a narrowed scope as well as smaller document numbers significantly lower the precision. In both cases, the modifications to the Termolator improve precision.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

ArticlesDocument Classification

Similar Papers 제목 키워드 기반

Building from Scratch: A Multi-Agent Framework with Human-in-the-Loop for Multilingual Legal Terminology Mapping

2025-12-15 · Lingyi Meng, Maolin Liu, Hao Wang, Yilan Cheng 외 arxiv

Accurately mapping legal terminology across languages remains a significant challenge, especially for language pairs like Chinese and Japanese, which share a large number of homographs with different meanings. Existing r…

TermGPT: Multi-Level Contrastive Fine-Tuning for Terminology Adaptation in Legal and Financial Domain

2025-11-13 · Yidan Sun, Mengying Zhu, Feiyue Chen, Yangyang Wu 외 arxiv

Large language models (LLMs) have demonstrated impressive performance in text generation tasks; however, their embedding spaces often suffer from the isotropy problem, resulting in poor discrimination of domain-specific …

Contrastive LearningText Generation

LegalRelectra: Mixed-domain Language Modeling for Long-range Legal Text Comprehension

2022-12-16 · Wenyue Hua, Yuchen Zhang, Zhe Chen, Josie Li 외

The application of Natural Language Processing (NLP) to specialized domains, such as the law, has recently received a surge of interest. As many legal services rely on processing and analyzing large collections of docume…

Language ModelingLanguage ModellingReading Comprehension

Adding a Third Language to a Lexical Resource Describing Legal Terminology: the assignment of equivalents

2014-05-01 · LREC 2014 5 · Janine Pimentel

Challenges and Considerations in Annotating Legal Data: A Comprehensive Overview

2024-07-05 · Harshil Darji, Jelena Mitrović, Michael Granitzer

The process of annotating data within the legal sector is filled with distinct challenges that differ from other fields, primarily due to the inherent complexities of legal language and documentation. The initial task us…