paper-with-me

Papers

Sampling-Frequency-Independent Universal Sound Separation

2023-09-22 · Tomohiko Nakamura, Kohei Yatabe

This paper proposes a universal sound separation (USS) method capable of handling untrained sampling frequencies (SFs). The USS aims at separating arbitrary sources of different types and can be the key technique to realize a source separator that can be universally used as a preprocessor for any downstream tasks. To realize a universal source separator, there are two essential properties: universalities with respect to source types and recording conditions. The former property has been studied in the USS literature, which has greatly increased the number of source types that can be handled by a single neural network. However, the latter property (e.g., SF) has received less attention despite its necessity. Since the SF varies widely depending on the downstream tasks, the universal source separator must handle a wide variety of SFs. In this paper, to encompass the two properties, we propose an SF-independent (SFI) extension of a computationally efficient USS network, SuDoRM-RF. The proposed network uses our previously proposed SFI convolutional layers, which can handle various SFs by generating convolutional kernels in accordance with an input SF. Experiments show that signal resampling can degrade the USS performance and the proposed method works more consistently than signal-resampling-based methods for various SFs.

📄 PDF Abstract BibTeX arXiv:2309.12581

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification

2024-09-19 · Dongheon Lee, Jung-Woo Choi

This paper presents a framework for universal sound separation and polyphonic audio classification, addressing the challenges of separating and classifying individual sound sources in a multichannel mixture. The proposed…

Audio ClassificationClassificationMamba

What's All the FUSS About Free Universal Sound Separation Data?

2020-11-02 · Scott Wisdom, Hakan Erdogan, Daniel Ellis, Romain Serizel 외

We introduce the Free Universal Sound Separation (FUSS) dataset, a new corpus for experiments in separating mixtures of an unknown number of sounds from an open domain of sound types. The dataset consists of 23 hours of …

AllData Augmentation

Unleashing the Power of Natural Audio Featuring Multiple Sound Sources

2025-04-24 · Xize Cheng, Slytherin Wang, Zehan Wang, Rongjie Huang 외

Universal sound separation aims to extract clean audio tracks corresponding to distinct events from mixed audio, which is critical for artificial auditory perception. However, current methods heavily rely on artificially…

Improving Universal Sound Separation Using Sound Classification

2019-11-18 · Efthymios Tzinis, Scott Wisdom, John R. Hershey, Aren Jansen 외

Deep learning approaches have recently achieved impressive performance on both audio source separation and sound classification. Most audio source separation approaches focus only on separating sources belonging to a res…

Audio Source SeparationClassificationGeneral ClassificationSound Classification

CLIPSep: Learning Text-queried Sound Separation with Noisy Unlabeled Videos

2022-12-14 · Hao-Wen Dong, Naoya Takahashi, Yuki Mitsufuji, Julian McAuley 외

Recent years have seen progress beyond domain-specific sound separation for speech or music towards universal sound separation for arbitrary sounds. Prior work on universal sound separation has investigated separating a …