paper-with-me

Papers Audio Source Separation

“Audio Source Separation” 태그가 달린 논문 117편 · 필터 해제

On loss functions and evaluation metrics for music source separation

2022-02-16 · Enric Gusó, Jordi Pons, Santiago Pascual, Joan Serrà

We investigate which loss functions provide better separations via benchmarking an extensive set of those for music source separation. To that end, we first survey the most representative audio source separation losses w…

Audio Source SeparationBenchmarkingMusic Source Separation

Differentiable Digital Signal Processing Mixture Model for Synthesis Parameter Extraction from Mixture of Harmonic Sounds

2022-02-01 · Masaya Kawamura, Tomohiko Nakamura, Daichi Kitamura, Hiroshi Saruwatari 외

A differentiable digital signal processing (DDSP) autoencoder is a musical sound synthesizer that combines a deep neural network (DNN) and spectral modeling synthesis. It allows us to flexibly edit sounds by changing the…

Audio Source Separation

Unsupervised Music Source Separation Using Differentiable Parametric Source Models

2022-01-24 · Kilian Schulze-Forster, Gaël Richard, Liam Kelley, Clement S. J. Doire 외

Supervised deep learning approaches to underdetermined audio source separation achieve state-of-the-art performance but require a dataset of mixtures along with their corresponding isolated source signals. Such datasets …

Audio Source SeparationDeep LearningMusic Source SeparationVocal ensemble separation

Fish sounds: towards the evaluation of marine acoustic biodiversity through data-driven audio source separation

2022-01-13 · Michele Mancusi, Nicola Zonca, Emanuele Rodolà, Silvia Zuffi

The marine ecosystem is changing at an alarming rate, exhibiting biodiversity loss and the migration of tropical species to temperate basins. Monitoring the underwater environments and their inhabitants is of fundamental…

Audio Source Separation

Self-Supervised Beat Tracking in Musical Signals with Polyphonic Contrastive Learning

2022-01-05 · Dorian Desblancs

Annotating musical beats is a very long and tedious process. In order to combat this problem, we present a new self-supervised learning pretext task for beat tracking and downbeat estimation. This task makes use of Splee…

Audio Source SeparationBeat TrackingContrastive LearningSelf-Supervised Learning

Zero-shot Audio Source Separation through Query-based Learning from Weakly-labeled Data

2021-12-15 · Ke Chen, Xingjian Du, Bilei Zhu, Zejun Ma 외

Deep learning techniques for separating audio into different sound sources face several challenges. Standard architectures require training separate models for different types of audio sources. Although some universal se…

Audio Source SeparationAudio TaggingEvent DetectionSound Event Detection+1

Zero-shot Audio Source Separation through Query-based Learningfrom Weakly-labeled Data

2021-12-15 · AAAI 2021 12 · Ke Chen, Xingjian Du, Bilei Zhu, Zejun Ma 외

Deep learning techniques for separating audio into different sound sources face several challenges. Standard architectures require training separate models for different types of audio sources. Although some universal se…

Audio Source SeparationEvent DetectionSound Event DetectionZero-shot Generalization

Hybrid Neural Networks for On-device Directional Hearing

2021-12-11 · AAAI 2022 2 · Anran Wang, Maruchi Kim, Hao Zhang, Shyamnath Gollakota

On-device directional hearing requires audio source separation from a given direction while achieving stringent human-imperceptible latency requirements. While neural nets can achieve significantly better performance tha…

Audio Source SeparationCausal InferenceDirectional HearingReal-time Directional Hearing

Transfer Learning with Jukebox for Music Source Separation

2021-11-28 · W. Zai El Amri, O. Tautz, H. Ritter, A. Melnik

In this work, we demonstrate how a publicly available, pre-trained Jukebox model can be adapted for the problem of audio source separation from a single mixed audio channel. Our neural network architecture, which is usin…

Audio Source SeparationMusic Source SeparationTransfer Learning

Reduction of Subjective Listening Effort for TV Broadcast Signals with Recurrent Neural Networks

2021-11-02 · Nils L. Westhausen, Rainer Huber, Hannah Baumgartner, Ragini Sinha 외

Listening to the audio of TV broadcast signals can be challenging for hearing-impaired as well as normal-hearing listeners, especially when background sounds are prominent or too loud compared to the speech signal. This …

Audio Source SeparationSpeech Enhancement

Unsupervised Source Separation By Steering Pretrained Music Models

2021-10-25 · Ethan Manilow, Patrick O'Reilly, Prem Seetharaman, Bryan Pardo

We showcase an unsupervised method that repurposes deep models trained for music generation and music tagging for audio source separation, without any retraining. An audio generation model is conditioned on an input mixt…

Audio GenerationAudio Source SeparationMusic GenerationMusic Tagging+1

The Cocktail Fork Problem: Three-Stem Audio Separation for Real-World Soundtracks

2021-10-19 · Darius Petermann, Gordon Wichern, Zhong-Qiu Wang, Jonathan Le Roux

The cocktail party problem aims at isolating any source of interest within a complex acoustic scene, and has long inspired audio source separation research. Recent efforts have mainly focused on separating speech from no…

Audio Source Separation

Unsupervised Source Separation via Bayesian Inference in the Latent Domain

2021-10-11 · Michele Mancusi, Emilian Postolache, Giorgio Mariani, Marco Fumero 외

State of the art audio source separation models rely on supervised data-driven approaches, which can be expensive in terms of labeling resources. On the other hand, approaches for training these models without any direct…

Audio Source SeparationBayesian InferenceMusic Source Separation

Visual Scene Graphs for Audio Source Separation

2021-09-24 · ICCV 2021 10 · Moitreya Chatterjee, Jonathan Le Roux, Narendra Ahuja, Anoop Cherian

State-of-the-art approaches for visually-guided audio source separation typically assume sources that have characteristic sounds, such as musical instruments. These approaches often ignore the visual context of these sou…

Audio Source SeparationVisually Guided Sound Source Separation

Multi-Task Audio Source Separation

2021-07-14 · Lu Zhang, Chenxing Li, Feng Deng, Xiaorui Wang

The audio source separation tasks, such as speech enhancement, speech separation, and music source separation, have achieved impressive performance in recent studies. The powerful modeling capabilities of deep neural net…

Audio Source SeparationMulti-task Audio Source SeperationMusic Source SeparationSpeech Enhancement+1

Densely Connected Multi-Dilated Convolutional Networks for Dense Prediction Tasks

2021-06-19 · CVPR 2021 1 · Naoya Takahashi, Yuki Mitsufuji

Tasks that involve high-resolution dense prediction require a modeling of both local and global patterns in a large input field. Although the local and global structures often depend on each other and their simultane…

Audio Source SeparationSemantic Segmentation

Parallel and Flexible Sampling from Autoregressive Models via Langevin Dynamics

2021-05-17 · Vivek Jayaram, John Thickstun

This paper introduces an alternative approach to sampling from autoregressive models. Autoregressive models are typically sampled sequentially, according to the transition dynamics defined by the model. Instead, we propo…

Audio Source SeparationSuper-Resolution

Move2Hear: Active Audio-Visual Source Separation

2021-05-15 · ICCV 2021 10 · Sagnik Majumder, Ziad Al-Halah, Kristen Grauman

We introduce the active audio-visual source separation problem, where an agent must move intelligently in order to better isolate the sounds coming from an object of interest in its environment. The agent hears multiple …

Audio Source SeparationObject

Sampling-Frequency-Independent Audio Source Separation Using Convolution Layer Based on Impulse Invariant Method

2021-05-10 · Koichi Saito, Tomohiko Nakamura, Kohei Yatabe, Yuma Koizumi 외

Audio source separation is often used as preprocessing of various applications, and one of its ultimate goals is to construct a single versatile model capable of dealing with the varieties of audio signals. Since samplin…

Audio Source SeparationMusic Source Separation

MULTIMODAL ANALYSIS: Informed content estimation and audio source separation

2021-04-27 · Gabriel Meseguer-Brocal

This dissertation proposes the study of multimodal learning in the context of musical signals. Throughout, we focus on the interaction between audio signals and text information. Among the many text sources related to mu…

Audio Source Separation
← 이전 41–60 / 117 다음 →