paper-with-me

Papers

Meta-Learning in Audio and Speech Processing: An End to End Comprehensive Review

2024-08-19 · Athul Raimon, Shubha Masti, Shyam K Sateesh, Siyani Vengatagiri, Bhaskarjyoti Das

This survey overviews various meta-learning approaches used in audio and speech processing scenarios. Meta-learning is used where model performance needs to be maximized with minimum annotated samples, making it suitable for low-sample audio processing. Although the field has made some significant contributions, audio meta-learning still lacks the presence of comprehensive survey papers. We present a systematic review of meta-learning methodologies in audio processing. This includes audio-specific discussions on data augmentation, feature extraction, preprocessing techniques, meta-learners, task selection strategies and also presents important datasets in audio, together with crucial real-world use cases. Through this extensive review, we aim to provide valuable insights and identify future research directions in the intersection of meta-learning and audio processing.

📄 PDF Abstract BibTeX arXiv:2408.10330

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationMeta-LearningSurvey

Similar Papers 제목 키워드 기반

Learnings from curating a trustworthy, well-annotated, and useful dataset of disordered English speech

2024-09-13 · Pan-Pan Jiang, Jimmy Tobin, Katrin Tomanek, Robert L. MacDonald 외

Project Euphonia, a Google initiative, is dedicated to improving automatic speech recognition (ASR) of disordered speech. A central objective of the project is to create a large, high-quality, and diverse speech corpus. …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

A Comprehensive Review and Taxonomy of Audio-Visual Synchronization Techniques for Realistic Speech Animation

2024-07-24 · Jose Geraldo Fernandes, Sinval Nascimento, Daniel Dominguete, André Oliveira 외

In many applications, synchronizing audio with visuals is crucial, such as in creating graphic animations for films or games, translating movie audio into different languages, and developing metaverse applications. This …

Audio-Visual Synchronization

Audio Self-supervised Learning: A Survey

2022-03-02 · Shuo Liu, Adria Mallol-Ragolta, Emilia Parada-Cabeleiro, Kun Qian 외

Inspired by the humans' cognitive ability to generalise knowledge and skills, Self-Supervised Learning (SSL) targets at discovering general representations from large-scale data without requiring human annotations, which…

Self-Supervised LearningSurvey

A Survey on Audio Synthesis and Audio-Visual Multimodal Processing

2021-08-01 · Zhaofeng Shi

With the development of deep learning and artificial intelligence, audio synthesis has a pivotal role in the area of machine learning and shows strong applicability in the industry. Meanwhile, significant efforts have be…

Audio SynthesisMusic GenerationSurveytext-to-speech+1

Deep Learning for Audio Signal Processing

2019-04-30 · Hendrik Purwins, Bo Li, Tuomas Virtanen, Jan Schlüter 외

Given the recent surge in developments of deep learning, this article provides a review of the state-of-the-art deep learning techniques for audio signal processing. Speech, music, and environmental sound processing are …

Audio Signal ProcessingAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Deep Learning+5