paper-with-me

Papers

Continual self-training with bootstrapped remixing for speech enhancement

2021-10-19 · Efthymios Tzinis, Yossi Adi, Vamsi K. Ithapu, Buye Xu, Anurag Kumar

We propose RemixIT, a simple and novel self-supervised training method for speech enhancement. The proposed method is based on a continuously self-training scheme that overcomes limitations from previous studies including assumptions for the in-domain noise distribution and having access to clean target signals. Specifically, a separation teacher model is pre-trained on an out-of-domain dataset and is used to infer estimated target signals for a batch of in-domain mixtures. Next, we bootstrap the mixing process by generating artificial mixtures using permuted estimated clean and noise signals. Finally, the student model is trained using the permuted estimated sources as targets while we periodically update teacher's weights using the latest student model. Our experiments show that RemixIT outperforms several previous state-of-the-art self-supervised methods under multiple speech enhancement tasks. Additionally, RemixIT provides a seamless alternative for semi-supervised and unsupervised domain adaptation for speech enhancement tasks, while being general enough to be applied to any separation task and paired with any separation model.

📄 PDF Abstract BibTeX arXiv:2110.10103

Code (1)

etzinis/unsup_speech_enh_adaptation pytorch

Tasks

Domain AdaptationSpeech EnhancementUnsupervised Domain Adaptation

Similar Papers 제목 키워드 기반

RemixIT: Continual self-training of speech enhancement models via bootstrapped remixing

2022-02-17 · Efthymios Tzinis, Yossi Adi, Vamsi Krishna Ithapu, Buye Xu 외

We present RemixIT, a simple yet effective self-supervised method for training speech enhancement without the need of a single isolated in-domain speech nor a noise waveform. Our approach overcomes limitations of previou…

Domain AdaptationSpeech EnhancementUnsupervised Domain Adaptation

Self-Remixing: Unsupervised Speech Separation via Separation and Remixing

2022-11-18 · Kohei Saijo, Tetsuji Ogawa

We present Self-Remixing, a novel self-supervised speech separation method, which refines a pre-trained separation model in an unsupervised manner. The proposed method consists of a shuffler module and a solver module, a…

Domain AdaptationSemi-supervised Domain AdaptationSpeech Separation

Remixing-based Unsupervised Source Separation from Scratch

2023-09-01 · Kohei Saijo, Tetsuji Ogawa

We propose an unsupervised approach for training separation models from scratch using RemixIT and Self-Remixing, which are recently proposed self-supervised learning methods for refining pre-trained models. They first se…

Self-Supervised Learning

Listen, Chat, and Remix: Text-Guided Soundscape Remixing for Enhanced Auditory Experience

2024-02-06 · Xilin Jiang, Cong Han, Yinghao Aaron Li, Nima Mesgarani

In daily life, we encounter a variety of sounds, both desirable and undesirable, with limited control over their presence and volume. Our work introduces "Listen, Chat, and Remix" (LCR), a novel multimodal sound remixer …

Language ModelingLanguage ModellingLarge Language Model

Bootstrapped Self-Supervised Training with Monocular Video for Semantic Segmentation and Depth Estimation

2021-03-19 · Yihao Zhang, John J. Leonard

For a robot deployed in the world, it is desirable to have the ability of autonomous learning to improve its initial pre-set knowledge. We formalize this as a bootstrapped self-supervised learning problem where a system …

Depth EstimationSelf-Supervised LearningSemantic Segmentation