paper-with-me

Papers

Producing Unseen Morphological Variants in Statistical Machine Translation

2017-04-01 · EACL 2017 4 · Matthias Huck, Ale{\v{s}} Tamchyna, Ond{\v{r}}ej Bojar, Alex Fraser, er

Translating into morphologically rich languages is difficult. Although the coverage of lemmas may be reasonable, many morphological variants cannot be learned from the training data. We present a statistical translation system that is able to produce these inflected word forms. Different from most previous work, we do not separate morphological prediction from lexical choice into two consecutive steps. Our approach is novel in that it is integrated in decoding and takes advantage of context information from both the source language and the target language sides.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

Augmenting Statistical Machine Translation with Subword Translation of Out-of-Vocabulary Words

2018-08-16 · Nelson F. Liu, Jonathan May, Michael Pust, Kevin Knight

Most statistical machine translation systems cannot translate words that are unseen in the training data. However, humans can translate many classes of out-of-vocabulary (OOV) words (e.g., novel morphological variants, m…

Machine TranslationTranslation

Role of Morphology Injection in Statistical Machine Translation

2017-09-16 · Sreelekha. S, Pushpak Bhattacharyya

Phrase-based Statistical models are more commonly used as they perform optimally in terms of both, translation quality and complexity of the system. Hindi and in general all Indian languages are morphologically richer th…

Machine TranslationTranslation

More than Just Statistical Recurrence: Human and Machine Unsupervised Learning of Māori Word Segmentation across Morphological Processes

2024-03-21 · Ashvini Varatharaj, Simon Todd

Non-M\=aori-speaking New Zealanders (NMS)are able to segment M\=aori words in a highlysimilar way to fluent speakers (Panther et al.,2024). This ability is assumed to derive through the identification and extraction of s…

Improving the Performance of English-Tamil Statistical Machine Translation System using Source-Side Pre-Processing

2014-09-29 · M. Anand Kumar, V. Dhanalakshmi, K. P. Soman, V. Sharmiladevi

Machine Translation is one of the major oldest and the most active research area in Natural Language Processing. Currently, Statistical Machine Translation (SMT) dominates the Machine Translation research. Statistical Ma…

Machine TranslationSentenceTranslation

Morphological Priors for Probabilistic Neural Word Embeddings

2016-08-03 · EMNLP 2016 11 · Parminder Bhatia, Robert Guthrie, Jacob Eisenstein

Word embeddings allow natural language processing systems to share statistical information across related words. These embeddings are typically based on distributional statistics, making it difficult for them to generali…

Part-Of-Speech TaggingWord EmbeddingsWord Similarity