paper-with-me

홈 › Papers

Basque Speecon-like and Basque SpeechDat MDB-600: speech databases for the development of ASR technology for Basque

2014-05-01 · LREC 2014 5 · Igor Odriozola, Inma Hernaez, Mar{\'\i}a In{\'e}s Torres, Luis Javier Rodriguez-Fuentes, Mikel Penagarikano, Eva Navas

This paper introduces two databases specifically designed for the development of ASR technology for the Basque language: the Basque Speecon-like database and the Basque SpeechDat MDB-600 database. The former was recorded in an office environment according to the Speecon specifications, whereas the later was recorded through mobile telephones according to the SpeechDat specifications. Both databases were created under an initiative that the Basque Government started in 2005, a program called ADITU, which aimed at developing speech technologies for Basque. The databases belong to the Basque Government. A comprehensive description of both databases is provided in this work, highlighting the differences with regard to their corresponding standard specifications. The paper also presents several initial experimental results for both databases with the purpose of validating their usefulness for the development of speech recognition technology. Several applications already developed with the Basque Speecon-like database are also described. Authors aim to make these databases widely known to the community as well, and foster their use by other groups.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Adding the Basque Parliament Corpus to ParlaMint Project

2022-06-01 · ParlaCLARIN (LREC) 2022 6 · Jon Alkorta, Mikel Iruskieta Quintian

The aim of this work is to describe the colection created with transcript of the Basque parliamentary speeches. This corpus follows the constraints of the ParlaMint project. The Basque ParlaMint corpus consists of two ve…

BasqueParl: A Bilingual Corpus of Basque Parliamentary Transcriptions

2022-05-03 · LREC 2022 6 · Nayla Escribano, Jon Ander González, Julen Orbegozo-Terradillos, Ainara Larrondo-Ureta 외

Parliamentary transcripts provide a valuable resource to understand the reality and know about the most important facts that occur over time in our societies. Furthermore, the political debates captured in these transcri…

Versatile Speech Databases for High Quality Synthesis for Basque

2012-05-01 · LREC 2012 5 · I{\~n}aki Sainz, Daniel Erro, Eva Navas, Inma Hern{\'a}ez 외

This paper presents three new speech databases for standard Basque. They are designed primarily for corpus-based synthesis but each database has its specific purpose: 1) AhoSyn: high quality speech synthesis (recorded al…

Emotional Speech SynthesisSpeech SynthesisVocal Bursts Intensity PredictionVoice Conversion

A Singing Voice Database in Basque for Statistical Singing Synthesis of Bertsolaritza

2016-05-01 · LREC 2016 5 · Xabier Sarasola, Eva Navas, David Tavarez, Daniel Erro 외

This paper describes the characteristics and structure of a Basque singing voice database of bertsolaritza. Bertsolaritza is a popular singing style from Basque Country sung exclusively in Basque that is improvised and a…

Singing Voice Synthesis

Dealing with dialectal variation in the construction of the Basque historical corpus

2020-12-01 · VarDial (COLING) 2020 12 · Ainara Estarrona, Izaskun Etxeberria, Ricardo Etxepare, Manuel Padilla-Moyano 외

This paper analyses the challenge of working with dialectal variation when semi-automatically normalising and analysing historical Basque texts. This work is part of a more general ongoing project for the construction of…