paper-with-me

Papers

Korean-to-Chinese Machine Translation using Chinese Character as Pivot Clue

2019-11-25 · Jeonghyeok Park, Hai Zhao

Korean-Chinese is a low resource language pair, but Korean and Chinese have a lot in common in terms of vocabulary. Sino-Korean words, which can be converted into corresponding Chinese characters, account for more than fifty of the entire Korean vocabulary. Motivated by this, we propose a simple linguistically motivated solution to improve the performance of the Korean-to-Chinese neural machine translation model by using their common vocabulary. We adopt Chinese characters as a translation pivot by converting Sino-Korean words in Korean sentences to Chinese characters and then train the machine translation model with the converted Korean sentences as source sentences. The experimental results on Korean-to-Chinese translation demonstrate that the models with the proposed method improve translation quality up to 1.5 BLEU points in comparison to the baseline models.

📄 PDF Abstract BibTeX arXiv:1911.11008

Code (2)

tmtmaj/Korean-Chinese-parallel-dataset-for-machine-translation-task
tmtmaj/Korean-Chinese_dataset

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

A Chinese POS Decision Method Using Korean Translation Information

2015-11-08 · Son-Il Kwak, O-Chol Kown, Chang-Sin Kim, Yong-Il Pak 외

In this paper we propose a method that imitates a translation expert using the Korean translation information and analyse the performance. Korean is good at tagging than Chinese, so we can use this property in Chinese PO…

POSPOS TaggingTranslation

Incorporating translation quality estimation into Chinese-Korean neural machine translation

2021-08-01 · CCL 2021 8 · Li Feiyu, Zhao Yahui, Yang Feiyang, Cui Rongyi

“Exposure bias and poor translation diversity are two common problems in neural machine trans-lation (NMT) which are caused by the general of the teacher forcing strategy for training inthe NMT models. Moreover the NMT m…

Machine TranslationNMTTranslation

Japanese to English/Chinese/Korean Datasets for Translation Quality Estimation and Automatic Post-Editing

2017-11-01 · WS 2017 11 · Atsushi Fujita, Eiichiro Sumita

Aiming at facilitating the research on quality estimation (QE) and automatic post-editing (APE) of machine translation (MT) outputs, especially for those among Asian languages, we have created new datasets for Japanese t…

Automatic Post-EditingBenchmarkingMachine TranslationSentence+1

HERITAGE: An End-to-End Web Platform for Processing Korean Historical Documents in Hanja

2025-01-21 · Seyoung Song, Haneul Yoo, Jiho Jin, Kyunghyun Cho 외

While Korean historical documents are invaluable cultural heritage, understanding those documents requires in-depth Hanja expertise. Hanja is an ancient language used in Korea before the 20th century, whose characters we…

document understandingMachine Translationnamed-entity-recognitionNamed Entity Recognition+1

A Study on the Korean and Chinese Pronunciation of Chinese Characters and Learning Korean as a Second Language

2018-12-01 · PACLIC 2018 12 · Xiao Luo, Yike Yang, Jing Sun