paper-with-me

홈 › Papers

Robustness to Capitalization Errors in Named Entity Recognition

2019-11-13 · WS 2019 11 · Sravan Bodapati, Hyokun Yun, Yaser Al-Onaizan

Robustness to capitalization errors is a highly desirable characteristic of named entity recognizers, yet we find standard models for the task are surprisingly brittle to such noise. Existing methods to improve robustness to the noise completely discard given orthographic information, mwhich significantly degrades their performance on well-formed text. We propose a simple alternative approach based on data augmentation, which allows the model to \emph{learn} to utilize or ignore orthographic information depending on its usefulness in the context. It achieves competitive robustness to capitalization errors while making negligible compromise to its performance on well-formed text and significantly improving generalization power on noisy user-generated text. Our experiments clearly and consistently validate our claim across different types of machine learning models, languages, and dataset sizes.

📄 PDF Abstract BibTeX arXiv:1911.05241

Code (0)

등록된 구현이 없습니다.

Tasks

Data Augmentationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)

Similar Papers 제목 키워드 기반

Bridge-Language Capitalization Inference in Western Iranian: Sorani, Kurmanji, Zazaki, and Tajik

2016-05-01 · LREC 2016 5 · Patrick Littell, David R. Mortensen, Kartik Goyal, Chris Dyer 외

In Sorani Kurdish, one of the most useful orthographic features in named-entity recognition {--} capitalization {--} is absent, as the language{'}s Perso-Arabic script does not make a distinction between uppercase and lo…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)

A Survey of Named Entity Recognition in Assamese and other Indian Languages

2014-07-09 · Gitimoni Talukdar, Pranjal Protim Borah, Arup Baruah

Named Entity Recognition is always important when dealing with major Natural Language Processing tasks such as information extraction, question-answering, machine translation, document summarization etc so in this paper …

Document SummarizationMachine Translationnamed-entity-recognitionNamed Entity Recognition+3

Improving Vietnamese Named Entity Recognition from Speech Using Word Capitalization and Punctuation Recovery Models

2020-10-01 · Thai Binh Nguyen, Quang Minh Nguyen, Thi Thu Hien Nguyen, Quoc Truong Do 외

Studies on the Named Entity Recognition (NER) task have shown outstanding results that reach human parity on input texts with correct text formattings, such as with proper punctuation and capitalization. However, such co…

Language ModelingLanguage Modellingnamed-entity-recognitionNamed Entity Recognition+4

Experiments to Improve Named Entity Recognition on Turkish Tweets

2014-10-31 · WS 2014 4 · Dilek Küçük, Ralf Steinberger

Social media texts are significant information sources for several application areas including trend analysis, event monitoring, and opinion mining. Unfortunately, existing solutions for tasks such as named entity recogn…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)Opinion Mining

Improving Named Entity Recognition in Spoken Dialog Systems by Context and Speech Pattern Modeling

2021-07-01 · SIGDIAL (ACL) 2021 7 · Minh Nguyen, Zhou Yu

While named entity recognition (NER) from speech has been around as long as NER from written text has, the accuracy of NER from speech has generally been much lower than that of NER from text. The rise in popularity of s…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)named-entity-recognitionNamed Entity Recognition+4