paper-with-me

홈 › Papers

CAMÕES: A Comprehensive Automatic Speech Recognition Benchmark for European Portuguese

2025-08-27 · Carlos Carvalho, Francisco Teixeira, Catarina Botelho, Anna Pompili, Rubén Solera-Ureña, Sérgio Paulo, Mariana Julião, Thomas Rolland, John Mendonça, Diogo Pereira, Isabel Trancoso, Alberto Abad arxiv

Existing resources for Automatic Speech Recognition in Portuguese are mostly focused on Brazilian Portuguese, leaving European Portuguese (EP) and other varieties under-explored. To bridge this gap, we introduce CAMÕES, the first open framework for EP and other Portuguese varieties. It consists of (1) a comprehensive evaluation benchmark, including 46h of EP test data spanning multiple domains; and (2) a collection of state-of-the-art models. For the latter, we consider multiple foundation models, evaluating their zero-shot and fine-tuned performances, as well as E-Branchformer models trained from scratch. A curated set of 425h of EP was used for both fine-tuning and training. Our results show comparable performance for EP between fine-tuned foundation models and the E-Branchformer. Furthermore, the best-performing models achieve relative improvements above 35% WER, compared to the strongest zero-shot foundation model, establishing a new state-of-the-art for EP and other varieties.

📄 PDF Abstract BibTeX arXiv:2508.19721

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Recognition

Similar Papers 제목 키워드 기반

ELITR: European Live Translator

2020-11-01 · EAMT 2020 11 · Ondřej Bojar, Dominik Macháček, Sangeet Sagar, Otakar Smrž 외

ELITR (European Live Translator) project aims to create a speech translation system for simultaneous subtitling of conferences and online meetings targetting up to 43 languages. The technology is tested by the Supreme Au…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine Translationspeech-recognition+2

A Speech Test Set of Practice Business Presentations with Additional Relevant Texts

2019-08-02 · Dominik Macháček, Jonáš Kratochvíl, Tereza Vojtěchová, Ondřej Bojar

We present a test corpus of audio recordings and transcriptions of presentations of students' enterprises together with their slides and web-pages. The corpus is intended for evaluation of automatic speech recognition (A…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

HESITA(te) in Portuguese

2014-05-01 · LREC 2014 5 · C, Sara eias, Dirce Celorico, Jorge Proen{\c{c}}a 외

Hesitations, so-called disfluencies, are a characteristic of spontaneous speech, playing a primary role in its structure, reflecting aspects of the language production and the management of inter-communication. In this p…

Acoustic ModellingAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Management+3

KALAKA-3: a database for the recognition of spoken European languages on YouTube audios

2014-05-01 · LREC 2014 5 · Luis Javier Rodr{\'\i}guez-Fuentes, Mikel Penagarikano, Amparo Varona, Mireia Diez 외

This paper describes the main features of KALAKA-3, a speech database specifically designed for the development and evaluation of language recognition systems. The database provides TV broadcast speech for training, and …

A Semi-Automated Live Interlingual Communication Workflow Featuring Intralingual Respeaking: Evaluation and Benchmarking

2022-06-01 · LREC 2022 6 · Tomasz Korybski, Elena Davitti, Constantin Orasan, Sabine Braun

In this paper, we present a semi-automated workflow for live interlingual speech-to-text communication which seeks to reduce the shortcomings of existing ASR systems: a human respeaker works with a speaker-dependent spee…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)BenchmarkingMachine Translation+4