paper-with-me

Papers

Corpus Creation and Analysis for Named Entity Recognition in Telugu-English Code-Mixed Social Media Data

2019-07-01 · ACL 2019 7 · Vamshi Krishna Srirangam, Appidi Abhinav Reddy, Vinay Singh, Manish Shrivastava

Named Entity Recognition(NER) is one of the important tasks in Natural Language Processing(NLP) and also is a subtask of Information Extraction. In this paper we present our work on NER in Telugu-English code-mixed social media data. Code-Mixing, a progeny of multilingualism is a way in which multilingual people express themselves on social media by using linguistics units from different languages within a sentence or speech context. Entity Extraction from social media data such as tweets(twitter) is in general difficult due to its informal nature, code-mixed data further complicates the problem due to its informal, unstructured and incomplete information. We present a Telugu-English code-mixed corpus with the corresponding named entity tags. The named entities used to tag data are Person({}Per{'}), Organization({}Org{'}) and Location({`}Loc{'}). We experimented with the machine learning models Conditional Random Fields(CRFs), Decision Trees and BiLSTMs on our corpus which resulted in a F1-score of 0.96, 0.94 and 0.95 respectively.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Entity Extraction using GANnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NERSentenceTAG

Similar Papers 제목 키워드 기반

Domain Adaptation for Named Entity Recognition Using CRFs

2016-05-01 · LREC 2016 5 · Tian Tian, Marco Dinarelli, Isabelle Tellier, Pedro Dias Cardoso

In this paper we explain how we created a labelled corpus in English for a Named Entity Recognition (NER) task from multi-source and multi-domain data, for an industrial partner. We explain the specificities of this corp…

Domain Adaptationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+1

An Open Corpus for Named Entity Recognition in Historic Newspapers

2016-05-01 · LREC 2016 5 · Clemens Neudecker

The availability of openly available textual datasets ({``}corpora{''}) with highly accurate manual annotations ({``}gold standard{''}) of named entities (e.g. persons, locations, organizations, etc.) is crucial in the t…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)

Creation of Corpus and Analysis in Code-Mixed Kannada-English Social Media Data for POS Tagging

2020-12-01 · ICON 2020 12 · Abhinav Reddy Appidi, Vamshi Krishna Srirangam, Darsi Suhas, Manish Shrivastava

Part-of-Speech (POS) is one of the essential tasks for many Natural Language Processing (NLP) applications. There has been a significant amount of work done in POS tagging for resource-rich languages. POS tagging is an e…

coreference-resolutionCoreference Resolutionnamed-entity-recognitionNamed Entity Recognition+5

Named Entity Recognition on Turkish Tweets

2014-05-01 · LREC 2014 5 · Dilek K{\"u}{\c{c}}{\"u}k, Guillaume Jacquet, Ralf Steinberger

Various recent studies show that the performance of named entity recognition (NER) systems developed for well-formed text types drops significantly when applied to tweets. The only existing study for the highly inflected…

Articlesnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+1

Automatic Creation of Arabic Named Entity Annotated Corpus Using Wikipedia

2014-04-01 · EACL 2014 4 · Maha Althobaiti, Udo Kruschwitz, Massimo Poesio
Morphological AnalysisNamed Entity Recognition (NER)