Meta-Learning in Audio and Speech Processing: An End to End Comprehensive Review
This survey overviews various meta-learning approaches used in audio and speech processing scenarios. Meta-learning is used where model performance needs to be maximized with minimum annotated samples, making it suitable for low-sample audio processing. Although the field has made some significant contributions, audio meta-learning still lacks the presence of comprehensive survey papers. We present a systematic review of meta-learning methodologies in audio processing. This includes audio-specific discussions on data augmentation, feature extraction, preprocessing techniques, meta-learners, task selection strategies and also presents important datasets in audio, together with crucial real-world use cases. Through this extensive review, we aim to provide valuable insights and identify future research directions in the intersection of meta-learning and audio processing.
Code (0)
등록된 구현이 없습니다.
Tasks
Data AugmentationMeta-LearningSurveySimilar Papers 제목 키워드 기반
Learnings from curating a trustworthy, well-annotated, and useful dataset of disordered English speech
Project Euphonia, a Google initiative, is dedicated to improving automatic speech recognition (ASR) of disordered speech. A central objective of the project is to create a large, high-quality, and diverse speech corpus. …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1A Comprehensive Review and Taxonomy of Audio-Visual Synchronization Techniques for Realistic Speech Animation
In many applications, synchronizing audio with visuals is crucial, such as in creating graphic animations for films or games, translating movie audio into different languages, and developing metaverse applications. This …
Audio-Visual SynchronizationAudio Self-supervised Learning: A Survey
Inspired by the humans' cognitive ability to generalise knowledge and skills, Self-Supervised Learning (SSL) targets at discovering general representations from large-scale data without requiring human annotations, which…
Self-Supervised LearningSurveyA Survey on Audio Synthesis and Audio-Visual Multimodal Processing
With the development of deep learning and artificial intelligence, audio synthesis has a pivotal role in the area of machine learning and shows strong applicability in the industry. Meanwhile, significant efforts have be…
Audio SynthesisMusic GenerationSurveytext-to-speech+1Deep Learning for Audio Signal Processing
Given the recent surge in developments of deep learning, this article provides a review of the state-of-the-art deep learning techniques for audio signal processing. Speech, music, and environmental sound processing are …
Audio Signal ProcessingAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Deep Learning+5