Improving Translation Selection with Supersenses
Selecting appropriate translations for source words with multiple meanings still remains a challenge for statistical machine translation (SMT). One reason for this is that most SMT systems are not good at detecting the proper sense for a polysemic word when it appears in different contexts. In this paper, we adopt a supersense tagging method to annotate source words with coarse-grained ontological concepts. In order to enable the system to choose an appropriate translation for a word or phrase according to the annotated supersense of the word or phrase, we propose two translation models with supersense knowledge: a maximum entropy based model and a supersense embedding model. The effectiveness of our proposed models is validated on a large-scale English-to-Spanish translation task. Results indicate that our method can significantly improve translation quality via correctly conveying the meaning of the source language to the target language.
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationTranslationWord Sense DisambiguationSimilar Papers 제목 키워드 기반
Adpositional Supersenses for Mandarin Chinese
This study adapts Semantic Network of Adposition and Case Supersenses (SNACS) annotation to Mandarin Chinese and demonstrates that the same supersense categories are appropriate for Chinese adposition semantics. We annot…
Machine TranslationTranslationSuperNMT: Neural Machine Translation with Semantic Supersenses and Syntactic Supertags
In this paper we incorporate semantic supersensetags and syntactic supertag features into EN{--}FR and EN{--}DE factored NMT systems. In experiments on various test sets, we observe that such features (and particularly w…
Machine TranslationNamed Entity Recognition (NER)NMTPrepositional Phrase Attachment+2A corpus of preposition supersenses in English web reviews
We present the first corpus annotated with preposition supersenses, unlexicalized categories for semantic functions that can be marked by English prepositions (Schneider et al., 2015). That scheme improves upon its prede…
A Corpus of Adpositional Supersenses for Mandarin Chinese
Adpositions are frequent markers of semantic relations, but they are highly ambiguous and vary significantly from language to language. Moreover, there is a dearth of annotated corpora for investigating the cross-linguis…
TranslationHindi-Urdu Adposition and Case Supersenses v1.0
These are the guidelines for the application of SNACS (Semantic Network of Adposition and Case Supersenses; Schneider et al. 2018) to Modern Standard Hindi of Delhi. SNACS is an inventory of 50 supersenses (semantic labe…