Statistical Machine Transliteration Baselines for NEWS 2018
This paper reports the results of our trans-literation experiments conducted on NEWS 2018 Shared Task dataset. We focus on creating the baseline systems trained using two open-source, statistical transliteration tools, namely Sequitur and Moses. We discuss the pre-processing steps performed on this dataset for both the systems. We also provide a re-ranking system which uses top hypotheses from Sequitur and Moses to create a consolidated list of transliterations. The results obtained from each of these models can be used to present a good starting point for the participating teams.
Code (0)
등록된 구현이 없습니다.
Tasks
Information RetrievalMachine TranslationRe-RankingTransliterationSimilar Papers 제목 키워드 기반
A Deep Learning Based Approach to Transliteration
In this paper, we propose different architectures for language independent machine transliteration which is extremely important for natural language processing (NLP) applications. Though a number of statistical models fo…
Deep LearningInformation RetrievalMachine TranslationNMT+2NEWS 2018 Whitepaper
Transliteration is defined as phonetic translation of names across languages. Transliteration of Named Entities (NEs) is necessary in many applications, such as machine translation, corpus alignment, cross-language IR, i…
BenchmarkingMachine TranslationTranslationTransliterationReport of NEWS 2018 Named Entity Transliteration Shared Task
This report presents the results from the Named Entity Transliteration Shared Task conducted as part of The Seventh Named Entities Workshop (NEWS 2018) held at ACL 2018 in Melbourne, Australia. Similar to previous editio…
Information RetrievalMachine TranslationTransliteration