Sequence-to-sequence neural network models for transliteration
Transliteration is a key component of machine translation systems and software internationalization. This paper demonstrates that neural sequence-to-sequence models obtain state of the art or close to state of the art results on existing datasets. In an effort to make machine transliteration accessible, we open source a new Arabic to English transliteration dataset and our trained models.
Code (1)
Tasks
Machine TranslationTranslationTransliterationSimilar Papers 제목 키워드 기반
A Deep Learning Based Approach to Transliteration
In this paper, we propose different architectures for language independent machine transliteration which is extremely important for natural language processing (NLP) applications. Though a number of statistical models fo…
Deep LearningInformation RetrievalMachine TranslationNMT+2Neural Machine Transliteration: Preliminary Results
Machine transliteration is the process of automatically transforming the script of a word from a source language to a target language, while preserving pronunciation. Sequence to sequence learning has recently emerged as…
DecoderTransliterationLow-Resource Machine Transliteration Using Recurrent Neural Networks of Asian Languages
Grapheme-to-phoneme models are key components in automatic speech recognition and text-to-speech systems. With low-resource language pairs that do not have available and well-developed pronunciation lexicons, grapheme-to…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine Translationspeech-recognition+6Sinhala Transliteration: A Comparative Analysis Between Rule-based and Seq2Seq Approaches
Due to reasons of convenience and lack of tech literacy, transliteration (i.e., Romanizing native scripts instead of using localization tools) is eminently prevalent in the context of low-resource languages such as Sinha…
DecoderMachine TranslationNMTTransliteration