paper-with-me

홈 › Papers

Neutral TTS Female Voice Corpus in Brazilian Portuguese

2023-10-08 · XLI Simpósio Brasileiro de Telecomunicações e Processamento de Sinais (SBrT2023) 2023 10 · Pedro H. L. Leite, Edmundo Hoyle, Álvaro Antelo, Luiz F. Kruszielski, Luiz W. P. Biscainho

This paper introduces a new dataset designed to address the limitations in high-quality, diverse and representative datasets for training text-to-speech (TTS) models, specifically for female voices in Brazilian Portuguese. The dataset features a female voice recorded in a professional and controlled environment with neutral emotion and comprises more than 20 hours of recordings. The goal is to facilitate transfer learning and enable the development of more natural-sounding, high-quality, and gender-balanced TTS systems. Alongside the dataset, gender-aware voice transfer experiments are performed to understand the impact of utilizing gender-specific pretrained models for speech synthesis. The results obtained show that same-gender voice transfer yields better speech similarity and intelligibility when compared to cross-gender transfer, emphasizing the importance of gender-aware training procedures and highlighting the need for balanced gender data.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Synthesistext-to-speechText to SpeechTransfer Learning

Similar Papers 제목 키워드 기반

A Corpus of Neutral Voice Speech in Brazilian Portuguese

2021-05-21 · International Conference on Computational Processing of the Portuguese Language 2021 5 · Pedro H. L. Leite, Edmundo Hoyle, Álvaro Antelo, Luiz F. Kruszielski 외

This work presents a new database containing high sampling rate recordings of a single male speaker reading sentences in Brazilian Portuguese with neutral voice, along with the corresponding text corpus. Intended for syn…

Speech Synthesistext-to-speechText to Speech

Pretrained audio neural networks for Speech emotion recognition in Portuguese

2022-10-26 · Marcelo Matheus Gauy, Marcelo Finger

The goal of speech emotion recognition (SER) is to identify the emotional aspects of speech. The SER challenge for Brazilian Portuguese speech was proposed with short snippets of Portuguese which are classified as neutra…

Data AugmentationEmotion RecognitionSpeech Emotion RecognitionTransfer Learning

CORAA: a large corpus of spontaneous and prepared speech manually validated for speech recognition in Brazilian Portuguese

2021-10-14 · Arnaldo Candido Junior, Edresson Casanova, Anderson Soares, Frederico Santos de Oliveira 외

Automatic Speech recognition (ASR) is a complex and challenging task. In recent years, there have been significant advances in the area. In particular, for the Brazilian Portuguese (BP) language, there were about 376 hou…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Building The First English-Brazilian Portuguese Corpus for Automatic Post-Editing

2020-12-01 · COLING 2020 8 · Felipe Almeida Costa, Thiago castro Ferreira, Adriana Pagano, Wagner Meira

This paper introduces the first corpus for Automatic Post-Editing of English and a low-resource language, Brazilian Portuguese. The source English texts were extracted from the WebNLG corpus and automatically translated …

Automatic Post-EditingMachine TranslationTranslation

Propbank-Br: a Brazilian Treebank annotated with semantic role labels

2012-05-01 · LREC 2012 5 · Magali Sanches Duran, S Alu{\'\i}sio, ra Maria

This paper reports the annotation of a Brazilian Portuguese Treebank with semantic role labels following Propbank guidelines. A different language and a different parser output impact the task and require some decisions …

Machine TranslationQuestion AnsweringSemantic Role Labeling