Tools and resources for Romanian text-to-speech and speech-to-text applications
In this paper we introduce a set of resources and tools aimed at providing support for natural language processing, text-to-speech synthesis and speech recognition for Romanian. While the tools are general purpose and can be used for any language (we successfully trained our system for more than 50 languages and participated in the Universal Dependencies Shared Task), the resources are only relevant for Romanian language processing.
Code (4)
Tasks
speech-recognitionSpeech RecognitionSpeech SynthesisSpeech-to-Texttext-to-speechText to SpeechText-To-Speech SynthesisSimilar Papers 제목 키워드 기반
RSC: A Romanian Read Speech Corpus for Automatic Speech Recognition
Although many efforts have been made in the last decade to enhance the speech and language resources for Romanian, this language is still considered under-resourced. While for many other languages there are large speech …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech RecognitionMoNERo: a Biomedical Gold Standard Corpus for the Romanian Language
In an era when large amounts of data are generated daily in various fields, the biomedical field among others, linguistic resources can be exploited for various tasks of Natural Language Processing. Moreover, increasing …
Romanian Language Translation in the RELATE Platform
This paper presents the usage of the RELATE platform for translation tasks involving the Romanian language. Using this platform, it is possible to perform text and speech data translations, either for single documents or…
TranslationRSS-TOBI - A Prosodically Enhanced Romanian Speech Corpus
This paper introduces a recent development of a Romanian Speech corpus to include prosodic annotations of the speech data in the form of ToBI labels. We describe the methodology of determining the required pitch patterns…
Speech Synthesistext-to-speechText to SpeechText-To-Speech SynthesisUse Case: Romanian Language Resources in the LOD Paradigm
In this paper, we report on (i) the conversion of Romanian language resources to the Linked Open Data specifications and requirements, on (ii) their publication and (iii) interlinking with other language resources (for R…
Word Embeddings