paper-with-me

Papers

Dialect Adaptation and Data Augmentation for Low-Resource ASR: TalTech Systems for the MADASR 2023 Challenge

2023-10-26 · Tanel Alumäe, Jiaming Kong, Daniil Robnikov

This paper describes Tallinn University of Technology (TalTech) systems developed for the ASRU MADASR 2023 Challenge. The challenge focuses on automatic speech recognition of dialect-rich Indian languages with limited training audio and text data. TalTech participated in two tracks of the challenge: Track 1 that allowed using only the provided training data and Track 3 which allowed using additional audio data. In both tracks, we relied on wav2vec2.0 models. Our methodology diverges from the traditional procedure of finetuning pretrained wav2vec2.0 models in two key points: firstly, through the implementation of the aligned data augmentation technique to enhance the linguistic diversity of the training data, and secondly, via the application of deep prefix tuning for dialect adaptation of wav2vec2.0 models. In both tracks, our approach yielded significant improvements over the provided baselines, achieving the lowest word error rates across all participating teams.

📄 PDF Abstract BibTeX arXiv:2310.17448

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionData AugmentationDiversityspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Data-Augmentation-Based Dialectal Adaptation for LLMs

2024-04-11 · Fahim Faisal, Antonios Anastasopoulos

This report presents GMUNLP's participation to the Dialect-Copa shared task at VarDial 2024, which focuses on evaluating the commonsense reasoning capabilities of large language models (LLMs) on South Slavic micro-dialec…

Data AugmentationNatural Language Understanding

Doing More with Less: Data Augmentation for Sudanese Dialect Automatic Speech Recognition

2026-01-11 · Ayman Mansour arxiv

Although many Automatic Speech Recognition (ASR) systems have been developed for Modern Standard Arabic (MSA) and Dialectal Arabic (DA), few studies have focused on dialect-specific implementations, particularly for low-…

Speech RecognitionData AugmentationHoldout Set

Improving Low-Resource Dialect Classification Using Retrieval-based Voice Conversion

2025-07-04 · Lea Fischbach, Akbar Karimi, Caroline Kleen, Alfred Lameli 외 arxiv

Deep learning models for dialect identification are often limited by the scarcity of dialectal data. To address this challenge, we propose to use Retrieval-based Voice Conversion (RVC) as an effective data augmentation m…

Data AugmentationVoice Conversion

Low-resource Language Discrimination Towards Chinese Dialects with Transfer learning and Data Augmentation

2026-06-17 · Fan Xu, Yangjie Dan, Keyu Yan, Yong Ma 외 arxiv

Chinese dialects discrimination is a challenging natural language processing task due to scarce annotation resource. In this article, we develop a novel Chinese dialects discrimination framework with transfer learning an…

Speech RecognitionTransfer LearningData Augmentation

A Catalog of Basque Dialectal Resources: Online Collections and Standard-to-Dialectal Adaptations

2026-03-26 · Jaione Bengoetxea, Itziar Gonzalez-Dios, Rodrigo Agerri arxiv

Recent research on dialectal NLP has identified data scarcity as a primary limitation. To address this limitation, this paper presents a catalog of contemporary Basque dialectal data and resources, offering a systematic …

Natural Language Inference