paper-with-me

Papers

Unsupervised Context-Sensitive Spelling Correction of Clinical Free-Text with Word and Character N-Gram Embeddings

2017-08-01 · WS 2017 8 · Pieter Fivez, Simon {\v{S}}uster, Walter Daelemans

We present an unsupervised context-sensitive spelling correction method for clinical free-text that uses word and character n-gram embeddings. Our method generates misspelling replacement candidates and ranks them according to their semantic fit, by calculating a weighted cosine similarity between the vectorized representation of a candidate and the misspelling context. We greatly outperform two baseline off-the-shelf spelling correction tools on a manually annotated MIMIC-III test set, and counter the frequency bias of an optimized noisy channel model, showing that neural embeddings can be successfully exploited to include context-awareness in a spelling correction model.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Spelling Correction

Similar Papers 제목 키워드 기반

Unsupervised Context-Sensitive Spelling Correction of English and Dutch Clinical Free-Text with Word and Character N-Gram Embeddings

2017-10-19 · Pieter Fivez, Simon Šuster, Walter Daelemans

We present an unsupervised context-sensitive spelling correction method for clinical free-text that uses word and character n-gram embeddings. Our method generates misspelling replacement candidates and ranks them accord…

Spelling Correction

Context-Sensitive Malicious Spelling Error Correction

2019-01-23 · Hongyu Gong, Yuchen Li, Suma Bhat, Pramod Viswanath

Misspelled words of the malicious kind work by changing specific keywords and are intended to thwart existing automated applications for cyber-environment control such as harassing content detection on the Internet and e…

Spam detectionSpelling CorrectionWord Embeddings

Vartani Spellcheck -- Automatic Context-Sensitive Spelling Correction of OCR-generated Hindi Text Using BERT and Levenshtein Distance

2020-12-14 · Aditya Pal, Abhijit Mustafi

Traditional Optical Character Recognition (OCR) systems that generate text of highly inflectional Indic languages like Hindi tend to suffer from poor accuracy due to a wide alphabet set, compound characters and difficult…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+3

Arabisc: Context-Sensitive Neural Spelling Checker

2020-12-01 · Yasmin Moslem, Rejwanul Haque, Andy Way

Traditional statistical approaches to spelling correction usually consist of two consecutive processes — error detection and correction — and they are generally computationally intensive. Current state-of-the-art neural …

Language ModellingSentenceSpelling Correction

Similarity-Based Unsupervised Spelling Correction Using BioWordVec: Development and Usability Study of Bacterial Culture and Antimicrobial Susceptibility Reports

2021-02-22 · JMIR Medical Informatics 2021 2 · Taehyeong Kim, Sung Won Han, Minji Kang, Se Ha Lee 외

Background: Existing bacterial culture test results for infectious diseases are written in unrefined text, resulting in many problems, including typographical errors and stop words. Effective spelling correction process…

Cultural Vocal Bursts Intensity PredictionSpelling Correction