paper-with-me

Papers

Open-Source High Quality Speech Datasets for Basque, Catalan and Galician

2020-05-01 · LREC 2020 5 · Oddur Kjartansson, Alex Gutkin, er, Alena Butryna, Isin Demirsahin, Clara Rivera

This paper introduces new open speech datasets for three of the languages of Spain: Basque, Catalan and Galician. Catalan is furthermore the official language of the Principality of Andorra. The datasets consist of high-quality multi-speaker recordings of the three languages along with the associated transcriptions. The resulting corpora include over 33 hours of crowd-sourced recordings of 132 male and female native speakers. The recording scripts also include material for elicitation of global and local place names, personal and business names. The datasets are released under a permissive license and are available for free download for commercial, academic and personal use. The high-quality annotated speech datasets described in this paper can be used to, among other things, build text-to-speech systems, serve as adaptation data in automatic speech recognition and provide useful phonetic and phonological insights in corpus linguistics.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognitiontext-to-speechText to SpeechVocal Bursts Intensity Prediction

Similar Papers 제목 키워드 기반

A Toolbox for Construction and Analysis of Speech Datasets

2021-04-11 · Evelina Bakhturina, Vitaly Lavrukhin, Boris Ginsburg

Automatic Speech Recognition and Text-to-Speech systems are primarily trained in a supervised fashion and require high-quality, accurately labeled speech datasets. In this work, we examine common problems with speech dat…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+2

Vietnamese Automatic Speech Recognition: A Revisit

2026-03-16 · Thi Vu, Linh The Nguyen, Dat Quoc Nguyen arxiv

Automatic Speech Recognition (ASR) performance is heavily dependent on the availability of large-scale, high-quality datasets. For low-resource languages, existing open-source ASR datasets often suffer from insufficient …

Speech Recognition

OpenACE: An Open Benchmark for Evaluating Audio Coding Performance

2024-09-12 · Jozef Coldenhoff, Niclas Granqvist, Milos Cernak

Audio and speech coding lack unified evaluation and open-source testing. Many candidate systems were evaluated on proprietary, non-reproducible, or small data, and machine learning-based codecs are often tested on datase…

OpenBibleTTS: Large-Scale Speech Resources and TTS Models for Low-Resource Languages

2026-06-08 · David Guzmán, Luel Hagos Beyene, Jesujoba Oluwadara Alabi, Yejin Jeon 외 arxiv

Recent advances in neural text-to-speech (TTS) and multilingual speech generation have substantially improved synthetic speech quality, yet these gains remain unevenly distributed across the world's languages. Existing m…

Speech Synthesis

DiffSSD: A Diffusion-Based Dataset For Speech Forensics

2024-09-19 · Kratika Bhagtani, Amit Kumar Singh Yadav, Paolo Bestagini, Edward J. Delp

Diffusion-based speech generators are ubiquitous. These methods can generate very high quality synthetic speech and several recent incidents report their malicious use. To counter such misuse, synthetic speech detectors …