Language Identification and Named Entity Recognition in Hinglish Code Mixed Tweets
While growing code-mixed content on Online Social Networks(OSN) provides a fertile ground for studying various aspects of code-mixing, the lack of automated text analysis tools render such studies challenging. To meet this challenge, a family of tools for analyzing code-mixed data such as language identifiers, parts-of-speech (POS) taggers, chunkers have been developed. Named Entity Recognition (NER) is an important text analysis task which is not only informative by itself, but is also needed for downstream NLP tasks such as semantic role labeling. In this work, we present an exploration of automatic NER of code-mixed data. We compare our method with existing off-the-shelf NER tools for social media content,and find that our systems outperforms the best baseline by 33.18 {\%} (F1 score).
Code (0)
등록된 구현이 없습니다.
Tasks
Abuse DetectionChunkingLanguage Identificationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NERPOSSemantic Role LabelingSimilar Papers 제목 키워드 기반
Comparative Study of Pre-Trained BERT and Large Language Models for Code-Mixed Named Entity Recognition
Named Entity Recognition (NER) in code-mixed text, particularly Hindi-English (Hinglish), presents unique challenges due to informal structure, transliteration, and frequent language switching. This study conducts a comp…
ANEC: An Amharic Named Entity Corpus and Transformer Based Recognizer
Named Entity Recognition is an information extraction task that serves as a preprocessing step for other natural language processing tasks, such as machine translation, information retrieval, and question answering. Name…
imbalanced classificationInformation RetrievalMachine Translationnamed-entity-recognition+4"Translation can't change a name": Using Multilingual Data for Named Entity Recognition
Named Entities (NEs) are often written with no orthographic changes across different languages that share a common alphabet. We show that this can be leveraged so as to improve named entity recognition (NER) by using uns…
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+1Chemical Identification and Indexing in PubMed Articles via BERT and Text-to-Text Approaches
The Biocreative VII Track-2 challenge consists of named entity recognition, entity-linking (or entity-normalization), and topic indexing tasks -- with entities and topics limited to chemicals for this challenge. Named en…
ArticlesChemical IndexingEntity LinkingMetric Learning+5Boundary identification of events in clinical named entity recognition
The problem of named entity recognition in the medical/clinical domain has gained increasing attention do to its vital role in a wide range of clinical decision support applications. The identification of complete and co…
General Classificationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)