paper-with-me

홈 › Papers

A Neural Topic-Attention Model for Medical Term Abbreviation Disambiguation

2019-10-30 · Irene Li, Michihiro Yasunaga, Muhammed Yavuz Nuzumlali, Cesar Caraballo, Shiwani Mahajan, Harlan Krumholz, Dragomir Radev

Automated analysis of clinical notes is attracting increasing attention. However, there has not been much work on medical term abbreviation disambiguation. Such abbreviations are abundant, and highly ambiguous, in clinical documents. One of the main obstacles is the lack of large scale, balance labeled data sets. To address the issue, we propose a few-shot learning approach to take advantage of limited labeled data. Specifically, a neural topic-attention model is applied to learn improved contextualized sentence representations for medical term abbreviation disambiguation. Another vital issue is that the existing scarce annotations are noisy and missing. We re-examine and correct an existing dataset for training and collect a test set to evaluate the models fairly especially for rare senses. We train our model on the training set which contains 30 abbreviation terms as categories (on average, 479 samples and 3.24 classes in each term) selected from a public abbreviation disambiguation dataset, and then test on a manually-created balanced dataset (each class in each term has 15 samples). We show that enhancing the sentence representation with topic information improves the performance on small-scale unbalanced training datasets by a large margin, compared to a number of baseline models.

📄 PDF Abstract BibTeX arXiv:1910.14076

Code (1)

IreneZihuiLi/TopicAttentionMedicalAD 공식 구현 pytorch

Tasks

Few-Shot LearningSentence

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Training without training data: Improving the generalizability of automated medical abbreviation disambiguation

2019-12-12 · Marta Skreta, Aryan Arbabi, Jixuan Wang, Michael Brudno

Abbreviation disambiguation is important for automated clinical note processing due to the frequent use of abbreviations in clinical settings. Current models for automated abbreviation disambiguation are restricted by th…

Data Augmentation

Token Classification for Disambiguating Medical Abbreviations

2022-10-05 · Mucahit Cevik, Sanaz Mohammad Jafari, Mitchell Myers, Savas Yildirim

Abbreviations are unavoidable yet critical parts of the medical text. Using abbreviations, especially in clinical patient notes, can save time and space, protect sensitive information, and help avoid repetitions. However…

Classificationtext-classificationText Classificationtoken-classification+1

MeDAL: Medical Abbreviation Disambiguation Dataset for Natural Language Understanding Pretraining

2020-12-27 · EMNLP (ClinicalNLP) 2020 11 · Zhi Wen, Xing Han Lu, Siva Reddy

One of the biggest challenges that prohibit the use of many current NLP methods in clinical settings is the availability of public datasets. In this work, we present MeDAL, a large medical text dataset curated for abbrev…

Mortality PredictionNatural Language Understanding

Abbreviation Explorer - an interactive system for pre-evaluation of Unsupervised Abbreviation Disambiguation

2019-06-01 · NAACL 2019 6 · Manuel R. Ciosici, Ira Assent

We present Abbreviation Explorer, a system that supports interactive exploration of abbreviations that are challenging for Unsupervised Abbreviation Disambiguation (UAD). Abbreviation Explorer helps to identify long-form…

Unsupervised Abbreviation Detection in Clinical Narratives

2016-12-01 · WS 2016 12 · Markus Kreuzthaler, Michel Oleynik, Alex Avian, er 외

Clinical narratives in electronic health record systems are a rich resource of patient-based information. They constitute an ongoing challenge for natural language processing, due to their high compactness and abundance …

Feature EngineeringSentence