paper-with-me

Papers

Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection

2023-12-27 · Jinbo Hu, Yin Cao, Ming Wu, Qiuqiang Kong, Feiran Yang, Mark D. Plumbley, Jun Yang

Environment shifts and conflicts present significant challenges for learning-based sound event localization and detection (SELD) methods. SELD systems, when trained in particular acoustic settings, often show restricted generalization capabilities for diverse acoustic environments. Furthermore, obtaining annotated samples for spatial sound events is notably costly. Deploying a SELD system in a new environment requires extensive time for re-training and fine-tuning. To overcome these challenges, we propose environment-adaptive Meta-SELD, designed for efficient adaptation to new environments using minimal data. Our method specifically utilizes computationally synthesized spatial data and employs Model-Agnostic Meta-Learning (MAML) on a pre-trained, environment-independent model. The method then utilizes fast adaptation to unseen real-world environments using limited samples from the respective environments. Inspired by the Learning-to-Forget approach, we introduce the concept of selective memory as a strategy for resolving conflicts across environments. This approach involves selectively memorizing target-environment-relevant information and adapting to the new environments through the selective attenuation of model parameters. In addition, we introduce environment representations to characterize different acoustic settings, enhancing the adaptability of our attenuation approach to various environments. We evaluate our proposed method on the development set of the Sony-TAu Realistic Spatial Soundscapes 2023 (STARSS23) dataset and computationally synthesized scenes. Experimental results demonstrate the superior performance of the proposed method compared to conventional supervised learning methods, particularly in localization.

📄 PDF Abstract BibTeX arXiv:2312.16422

Code (1)

jinbo-hu/seld-data-generator 공식 구현

Tasks

Meta-LearningSound Event Localization and Detection

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

META-SELD: Meta-Learning for Fast Adaptation to the new environment in Sound Event Localization and Detection

2023-08-17 · Jinbo Hu, Yin Cao, Ming Wu, Feiran Yang 외

For learning-based sound event localization and detection (SELD) methods, different acoustic environments in the training and test sets may result in large performance differences in the validation and evaluation stages.…

Meta-LearningSound Event Localization and Detection

SelectTSL: Prompt-Guided Selective Target Sound Localization in Complex Scenarios

2026-07-02 · Ziyang Jiang, Yu Chen, Zexu Pan, Xinyuan Qian 외 arxiv

Humans can selectively attend to a target sound and estimate its direction in complex scenarios, whereas such selective localization remains challenging for current deep learning-based systems. Sound source localization …

Sound Source LocalizationTarget Sound Extraction

Weakly Supervised Context Encoder using DICOM metadata in Ultrasound Imaging

2020-03-20 · Szu-Yeu Hu, Shuhang Wang, Wei-Hung Weng, JingChao Wang 외

Modern deep learning algorithms geared towards clinical adaption rely on a significant amount of high fidelity labeled data. Low-resource settings pose challenges like acquiring high fidelity data and becomes the bottlen…

Learning How to Remember: A Meta-Cognitive Management Method for Structured and Transferable Agent Memory

2026-01-12 · Sirui Liang, Pengfei Cao, Jian Zhao, Wenhao Teng 외 arxiv

Large language model (LLM) agents increasingly rely on accumulated memory to solve long-horizon decision-making tasks. However, most existing approaches store memory in fixed representations and reuse it at a single or i…

The Spatial Selective Auditory Attention of Cochlear Implant Users in Different Conversational Sound Levels

2021-03-03 · Sara Akbarzadeh, Sungmin Lee, Chin-Tuan Tan

In multi speakers environments, cochlear implant (CI) users may attend to a target sound source in a different manner from the normal hearing (NH) individuals during a conversation. This study attempted to investigate th…

EEGElectroencephalogram (EEG)speech-recognitionSpeech Recognition