paper-with-me

Papers

ASR data augmentation in low-resource settings using cross-lingual multi-speaker TTS and cross-lingual voice conversion

2022-03-29 · Edresson Casanova, Christopher Shulby, Alexander Korolev, Arnaldo Candido Junior, Anderson da Silva Soares, Sandra Aluísio, Moacir Antonelli Ponti

We explore cross-lingual multi-speaker speech synthesis and cross-lingual voice conversion applied to data augmentation for automatic speech recognition (ASR) systems in low/medium-resource scenarios. Through extensive experiments, we show that our approach permits the application of speech synthesis and voice conversion to improve ASR systems using only one target-language speaker during model training. We also managed to close the gap between ASR models trained with synthesized versus human speech compared to other works that use many speakers. Finally, we show that it is possible to obtain promising ASR training results with our data augmentation method using only a single real speaker in a target language.

📄 PDF Abstract BibTeX arXiv:2204.00618

Code (1)

edresson/wav2vec-wrapper 공식 구현 pytorch

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognitionSpeech RecognitionSpeech SynthesisVoice Conversion

Similar Papers 제목 키워드 기반

Cross-lingual Back-Parsing: Utterance Synthesis from Meaning Representation for Zero-Resource Semantic Parsing

2024-10-01 · Deokhyung Kang, Seonjeong Hwang, Yunsu Kim, Gary Geunbae Lee

Recent efforts have aimed to utilize multilingual pretrained language models (mPLMs) to extend semantic parsing (SP) across multiple languages without requiring extensive annotations. However, achieving zero-shot cross-l…

Cross-Lingual TransferData AugmentationSemantic ParsingZero-Shot Cross-Lingual Transfer

Generalized Data Augmentation for Low-Resource Translation

2019-06-10 · ACL 2019 7 · Mengzhou Xia, Xiang Kong, Antonios Anastasopoulos, Graham Neubig

Translation to or from low-resource languages LRLs poses challenges for machine translation in terms of both adequacy and fluency. Data augmentation utilizing large amounts of monolingual data is regarded as an effective…

Data AugmentationMachine TranslationTranslationUnsupervised Machine Translation

ACLM: A Selective-Denoising based Generative Data Augmentation Approach for Low-Resource Complex NER

2023-06-01 · Sreyan Ghosh, Utkarsh Tyagi, Manan Suri, Sonal Kumar 외

Complex Named Entity Recognition (NER) is the task of detecting linguistically complex named entities in low-context text. In this paper, we present ACLM Attention-map aware keyword selection for Conditional Language Mod…

Data AugmentationDenoisingLanguage Modellingnamed-entity-recognition+4

CLASP: Few-Shot Cross-Lingual Data Augmentation for Semantic Parsing

2022-10-13 · Andy Rosenbaum, Saleh Soltan, Wael Hamza, Amir Saffari 외

A bottleneck to developing Semantic Parsing (SP) models is the need for a large volume of human-labeled training data. Given the complexity and cost of human annotation for SP, labeled data is often scarce, particularly …

Data AugmentationSemantic Parsing

Learning Cross-lingual Mappings for Data Augmentation to Improve Low-Resource Speech Recognition

2023-06-14 · Muhammad Umar Farooq, Thomas Hain

Exploiting cross-lingual resources is an effective way to compensate for data scarcity of low resource languages. Recently, a novel multilingual model fusion technique has been proposed where a model is trained to learn …

Data Augmentationspeech-recognitionSpeech RecognitionTransliteration