Turkish Named Entity Recognition: A Survey and Comparative Analysis
Named entity recognition is a challenging task that has been widely studied in English. Although there are some efforts for named entity recognition in Turkish language, the reported results are limited to particular datasets and models. Moreover, there is a lack of comparative analysis for named entity recognition in Turkish. In this study, we contribute to the literature in three folds. First, we provide an up-to-date short survey on Turkish named entity recognition studies. Second, we compare state-of-the-art named entity recognition models on various Turkish datasets that we can access to. Lastly, we analyze a set of linguistic processing steps that would affect the performance of Turkish named entity recognition.
Code (0)
등록된 구현이 없습니다.
Tasks
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)SurveySimilar Papers 제목 키워드 기반
Named Entity Recognition on Turkish Tweets
Various recent studies show that the performance of named entity recognition (NER) systems developed for well-formed text types drops significantly when applied to tweets. The only existing study for the highly inflected…
Articlesnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+1Experiments to Improve Named Entity Recognition on Turkish Tweets
Social media texts are significant information sources for several application areas including trend analysis, event monitoring, and opinion mining. Unfortunately, existing solutions for tasks such as named entity recogn…
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)Opinion MiningA Twitter Corpus for Named Entity Recognition in Turkish
This paper introduces a new Turkish Twitter Named Entity Recognition dataset. The dataset, which consists of 5000 tweets from a year-long period, was labeled by multiple annotators with a high agreement score. The datase…
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)Automatically Annotated Turkish Corpus for Named Entity Recognition and Text Categorization using Large-Scale Gazetteers
Turkish Wikipedia Named-Entity Recognition and Text Categorization (TWNERTC) dataset is a collection of automatically categorized and annotated sentences obtained from Wikipedia. We constructed large-scale gazetteers by …
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+1TR-SEQ: Named Entity Recognition Dataset for Turkish Search Engine Queries
Recognizing named entities in short search engine queries is a difficult task due to their weaker contextual information compared to long sentences. Standard named entity recognition (NER) systems that are trained on gra…
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER