paper-with-me

홈 › Papers

A Corpus of Neutral Voice Speech in Brazilian Portuguese

2021-05-21 · International Conference on Computational Processing of the Portuguese Language 2021 5 · Pedro H. L. Leite, Edmundo Hoyle, Álvaro Antelo, Luiz F. Kruszielski, Luiz W. P. Biscainho

This work presents a new database containing high sampling rate recordings of a single male speaker reading sentences in Brazilian Portuguese with neutral voice, along with the corresponding text corpus. Intended for synthesis and other speech-oriented applications, the dataset contains text scripts extracted from a popular Brazilian news TV program, read out loud by a trained individual in a controlled environment, resulting in roughly 20 h of audio data. The text was normalized in the recording process and special textual occurrences (e.g. acronyms, numbers, foreign names etc.) were replaced by their phonetic translation to a readable text in Portuguese. There are no noticeable accidental sounds and background noise has been kept to a minimum in all audio samples. To illustrate the potential benefits of having this data available, text-to-speech experiments were conducted using state-of-the-art models for speech synthesis (Tacotron 2 and Waveglow). As a result, we obtained intelligible and natural sounding voices from as few as 8 min of audio samples coming from an unseen target speaker, after having trained over our data; moreover, by increasing the target recording time to 75 min, we have noticeably improved accuracy in pronunciation.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Synthesistext-to-speechText to Speech

Similar Papers 제목 키워드 기반

Neutral TTS Female Voice Corpus in Brazilian Portuguese

2023-10-08 · XLI Simpósio Brasileiro de Telecomunicações e Processamento de Sinais (SBrT2023) 2023 10 · Pedro H. L. Leite, Edmundo Hoyle, Álvaro Antelo, Luiz F. Kruszielski 외

This paper introduces a new dataset designed to address the limitations in high-quality, diverse and representative datasets for training text-to-speech (TTS) models, specifically for female voices in Brazilian Portugues…

Speech Synthesistext-to-speechText to SpeechTransfer Learning

CORAA: a large corpus of spontaneous and prepared speech manually validated for speech recognition in Brazilian Portuguese

2021-10-14 · Arnaldo Candido Junior, Edresson Casanova, Anderson Soares, Frederico Santos de Oliveira 외

Automatic Speech recognition (ASR) is a complex and challenging task. In recent years, there have been significant advances in the area. In particular, for the Brazilian Portuguese (BP) language, there were about 376 hou…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Pretrained audio neural networks for Speech emotion recognition in Portuguese

2022-10-26 · Marcelo Matheus Gauy, Marcelo Finger

The goal of speech emotion recognition (SER) is to identify the emotional aspects of speech. The SER challenge for Brazilian Portuguese speech was proposed with short snippets of Portuguese which are classified as neutra…

Data AugmentationEmotion RecognitionSpeech Emotion RecognitionTransfer Learning

TTS-Portuguese Corpus: a corpus for speech synthesis in Brazilian Portuguese

2020-05-11 · Edresson Casanova, Arnaldo Candido Junior, Christopher Shulby, Frederico Santos de Oliveira 외

Speech provides a natural way for human-computer interaction. In particular, speech synthesis systems are popular in different applications, such as personal assistants, GPS applications, screen readers and accessibility…

DenoisingSpeech SynthesisTransfer Learning

Challenges in modality annotation in a Brazilian Portuguese Spontaneous Speech Corpus

2013-03-01 · WS 2013 3 · Luciana Beatriz Avila, Heliana Mello
Sentiment Analysis