paper-with-me

Papers

Experimental Evaluation and Development of a Silver-Standard for the MIMIC-III Clinical Coding Dataset

2020-06-12 · WS 2020 7 · Thomas Searle, Zina Ibrahim, Richard JB Dobson

Clinical coding is currently a labour-intensive, error-prone, but critical administrative process whereby hospital patient episodes are manually assigned codes by qualified staff from large, standardised taxonomic hierarchies of codes. Automating clinical coding has a long history in NLP research and has recently seen novel developments setting new state of the art results. A popular dataset used in this task is MIMIC-III, a large intensive care database that includes clinical free text notes and associated codes. We argue for the reconsideration of the validity MIMIC-III's assigned codes that are often treated as gold-standard, especially when MIMIC-III has not undergone secondary validation. This work presents an open-source, reproducible experimental methodology for assessing the validity of codes derived from EHR discharge summaries. We exemplify the methodology with MIMIC-III discharge summaries and show the most frequently assigned codes in MIMIC-III are under-coded up to 35%.

📄 PDF Abstract BibTeX arXiv:2006.07332

Code (1)

CogStack/MedCAT pytorch

Similar Papers 제목 키워드 기반

AugAbEx: Bridging Abstractive and Extractive Legal Summarization

2025-11-15 · Purnima Bindal, Vikas Kumar, Sagar Rathore, Vasudha Bhatnagar arxiv

Automatic summarization of legal judgments liberates law professionals from heavy cognitive burden due to the complexity of the language, context-sensitive legal jargon, and the length of the document. Caveats of abstrac…

Learning with Silver Standard Data for Zero-shot Relation Extraction

2022-11-25 · Tianyin Wang, Jianwei Wang, Ziqian Zeng

The superior performance of supervised relation extraction (RE) methods heavily relies on a large amount of gold standard data. Recent zero-shot relation extraction methods converted the RE task to other NLP tasks and us…

RelationRelation Extraction

On the use of Silver Standard Data for Zero-shot Classification Tasks in Information Extraction

2024-02-28 · Jianwei Wang, Tianyin Wang, Ziqian Zeng

The superior performance of supervised classification methods in the information extraction (IE) area heavily relies on a large amount of gold standard data. Recent zero-shot classification methods converted the task to …

ClassificationNatural Language InferenceRelation Classificationzero-shot-classification+2

Creating a silver standard for patent simplification

2023-10-24 · Silvia Casola, Alberto Lavelli, Horacio Saggion

Patents are legal documents that aim at protecting inventions on the one hand and at making technical knowledge circulate on the other. Their complex style -- a mix of legal, technical, and extremely vague language -- ma…

Information RetrievalRetrieval

Turning silver into gold: error-focused corpus reannotation with active learning

2019-09-01 · RANLP 2019 9 · Pierre Andr{\'e} M{\'e}nard, Antoine Mougeot

While high quality gold standard annotated corpora are crucial for most tasks in natural language processing, many annotated corpora published in recent years, created by annotators or tools, contains noisy annotations. …

Active LearningDocument ClassificationPart-Of-Speech Tagging