Mongolian Named Entity Recognition System with Rich Features
In this paper, we first build a manually annotated named entity corpus of Mongolian. Then, we propose three morphological processing methods and study comprehensive features, including syllable features, lexical features, context features, morphological features and semantic features in Mongolian named entity recognition. Moreover, we also evaluate the influence of word cluster features on the system and combine all features together eventually. The experimental result shows that segmenting each suffix into an individual token achieves better results than deleting suffixes or using the suffixes as feature. The system based on segmenting suffixes with all proposed features yields benchmark result of F-measure=84.65 on this corpus.
Code (0)
등록된 구현이 없습니다.
Tasks
Machine Translationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)Question AnsweringSimilar Papers 제목 키워드 기반
Attention-based BLSTM-CRF Architecture for Mongolian Named Entity Recognition
Feature-Rich Twitter Named Entity Recognition and Classification
Twitter named entity recognition is the process of identifying proper names and classifying them into some predefined labels/categories. The paper introduces a Twitter named entity system using a supervised machine learn…
ClassificationEntity Extraction using GANGeneral ClassificationMachine Translation+4Multilingual Slavic Named Entity Recognition
Named entity recognition, in particular for morphological rich languages, is challenging task due to the richness of inflected forms and ambiguity. This challenge is being addressed by SlavNER Shared Task. In this paper …
Language ModelingLanguage Modellingnamed-entity-recognitionNamed Entity Recognition+2MnTTS2: An Open-Source Multi-Speaker Mongolian Text-to-Speech Synthesis Dataset
Text-to-Speech (TTS) synthesis for low-resource languages is an attractive research issue in academia and industry nowadays. Mongolian is the official language of the Inner Mongolia Autonomous Region and a representative…
Speech Synthesistext-to-speechText to SpeechText-To-Speech SynthesisGHHT at CALCS 2018: Named Entity Recognition for Dialectal Arabic Using Neural Networks
This paper describes our system submission to the CALCS 2018 shared task on named entity recognition on code-switched data for the language variant pair of Modern Standard Arabic and Egyptian dialectal Arabic. We build a…
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)