paper-with-me

홈 › Papers

Jejueo Datasets for Machine Translation and Speech Synthesis

2019-11-27 · LREC 2020 5 · Kyubyong Park, Yo Joong Choe, Jiyeon Ham

Jejueo was classified as critically endangered by UNESCO in 2010. Although diverse efforts to revitalize it have been made, there have been few computational approaches. Motivated by this, we construct two new Jejueo datasets: Jejueo Interview Transcripts (JIT) and Jejueo Single Speaker Speech (JSS). The JIT dataset is a parallel corpus containing 170k+ Jejueo-Korean sentences, and the JSS dataset consists of 10k high-quality audio files recorded by a native Jejueo speaker and a transcript file. Subsequently, we build neural systems of machine translation and speech synthesis using them. All resources are publicly available via our GitHub repository. We hope that these datasets will attract interest of both language and machine learning communities.

📄 PDF Abstract BibTeX arXiv:1911.12071

Code (1)

kakaobrain/jejueo 공식 구현 tf

Tasks

Machine TranslationSpeech SynthesisTranslation

Similar Papers 제목 키워드 기반

Improving Jejueo-Korean Translation With Cross-Lingual Pretraining Using Japanese and Korean

2022-10-01 · WAT 2022 10 · Francis Zheng, Edison Marrese-Taylor, Yutaka Matsuo

Jejueo is a critically endangered language spoken on Jeju Island and is closely related to but mutually unintelligible with Korean. Parallel data between Jejueo and Korean is scarce, and translation between the two langu…

Machine TranslationTranslation

Preparing an Endangered Language for the Digital Age: The Case of Judeo-Spanish

2022-05-31 · EURALI (LREC) 2022 6 · Alp Öktem, Rodolfo Zevallos, Yasmin Moslem, Güneş Öztürk 외

We develop machine translation and speech synthesis systems to complement the efforts of revitalizing Judeo-Spanish, the exiled language of Sephardic Jews, which survived for centuries, but now faces the threat of extinc…

Machine TranslationSpeech Synthesistext-to-speechText to Speech+2

A Survey of Voice Translation Methodologies - Acoustic Dialect Decoder

2016-10-13 · Hans Krupakar, Keerthika Rajvel, Bharathi B, Angel Deborah S 외

Speech Translation has always been about giving source text or audio input and waiting for system to give translated output in desired form. In this paper, we present the Acoustic Dialect Decoder (ADD) - a voice to voice…

DecoderSentenceSpeech SynthesisSurvey+1

Simultaneous Speech-to-Speech Translation System with Neural Incremental ASR, MT, and TTS

2020-11-10 · Katsuhito Sudoh, Takatomo Kano, Sashi Novitasari, Tomoya Yanagita 외

This paper presents a newly developed, simultaneous neural speech-to-speech translation system and its evaluation. The system consists of three fully-incremental neural processing modules for automatic speech recognition…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine TranslationSimultaneous Speech-to-Speech Translation+8

German-Arabic Speech-to-Speech Translation for Psychiatric Diagnosis

2020-12-01 · COLING (WANLP) 2020 12 · Juan Hussain, Mohammed Mediani, Moritz Behr, M. Amin Cheragui 외

In this paper we present the natural language processing components of our German-Arabic speech-to-speech translation system which is being deployed in the context of interpretation during psychiatric, diagnostic intervi…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DecoderDiagnostic+7