paper-with-me

홈 › Papers

Joint Blind Room Acoustic Characterization From Speech And Music Signals Using Convolutional Recurrent Neural Networks

2020-10-21 · Paul Callens, Milos Cernak

Acoustic environment characterization opens doors for sound reproduction innovations, smart EQing, speech enhancement, hearing aids, and forensics. Reverberation time, clarity, and direct-to-reverberant ratio are acoustic parameters that have been defined to describe reverberant environments. They are closely related to speech intelligibility and sound quality. As explained in the ISO3382 standard, they can be derived from a room measurement called the Room Impulse Response (RIR). However, measuring RIRs requires specific equipment and intrusive sound to be played. The recent audio combined with machine learning suggests that one could estimate those parameters blindly using speech or music signals. We follow these advances and propose a robust end-to-end method to achieve blind joint acoustic parameter estimation using speech and/or music signals. Our results indicate that convolutional recurrent neural networks perform best for this task, and including music in training also helps improve inference from speech.

📄 PDF Abstract BibTeX arXiv:2010.11167

Code (0)

등록된 구현이 없습니다.

Tasks

parameter estimationRoom Impulse Response (RIR)Speech Enhancement

Similar Papers 제목 키워드 기반

MOSRA: Joint Mean Opinion Score and Room Acoustics Speech Quality Assessment

2022-04-04 · Karl El Hajal, Milos Cernak, Pablo Mainar

The acoustic environment can degrade speech quality during communication (e.g., video call, remote presentation, outside voice recording), and its impact is often unknown. Objective metrics for speech quality have proven…

Unsupervised Blind Joint Dereverberation and Room Acoustics Estimation with Diffusion Models

2024-08-14 · Jean-Marie Lemercier, Eloi Moliner, Simon Welker, Vesa Välimäki 외

This paper presents an unsupervised method for single-channel blind dereverberation and room impulse response (RIR) estimation, called BUDDy. The algorithm is rooted in Bayesian posterior sampling: it combines a likeliho…

Room Impulse Response (RIR)Speech Dereverberation

A Universal Deep Room Acoustics Estimator

2021-09-29 · Paula Sánchez López, Paul Callens, Milos Cernak

Speech audio quality is subject to degradation caused by an acoustic environment and isotropic ambient and point noises. The environment can lead to decreased speech intelligibility and loss of focus and attention by the…

Room Impulse Response (RIR)

BUDDy: Single-Channel Blind Unsupervised Dereverberation with Diffusion Models

2024-05-07 · Eloi Moliner, Jean-Marie Lemercier, Simon Welker, Timo Gerkmann 외

In this paper, we present an unsupervised single-channel method for joint blind dereverberation and room impulse response estimation, based on posterior sampling with diffusion models. We parameterize the reverberation o…

Blind Room Parameter Estimation Using Multiple-Multichannel Speech Recordings

2021-07-29 · Prerak Srivastava, Antoine Deleforge, Emmanuel Vincent

Knowing the geometrical and acoustical parameters of a room may benefit applications such as audio augmented reality, speech dereverberation or audio forensics. In this paper, we study the problem of jointly estimating t…

parameter estimationSpeech Dereverberation