chrF++: words helping character n-grams
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationSimilar Papers 제목 키워드 기반
BPE beyond Word Boundary: How NOT to use Multi Word Expressions in Neural Machine Translation
BPE tokenization merges characters into longer tokens by finding frequently occurring contiguous patterns within the word boundary. An intuitive relaxation would be to extend a BPE vocabulary with multi-word expressions …
Machine TranslationNMTTranslationCIC-FBK Approach to Native Language Identification
We present the CIC-FBK system, which took part in the Native Language Identification (NLI) Shared Task 2017. Our approach combines features commonly used in previous NLI research, i.e., word n-grams, lemma n-grams, part-…
General ClassificationLanguage IdentificationLEMMANative Language IdentificationNLIP_Lab-IITH Multilingual MT System for WAT24 MT Shared Task
This paper describes NLIP Lab's multilingual machine translation system for the WAT24 shared task on multilingual Indic MT task for 22 scheduled languages belonging to 4 language families. We explore pre-training for Ind…
Machine TranslationTranslationComplex Word Identification Using Character n-grams
This paper investigates the use of character n-gram frequencies for identifying complex words in English, German and Spanish texts. The approach is based on the assumption that complex words are likely to contain differe…
Complex Word IdentificationLexical SimplificationMachine TranslationText Classification