paper-with-me

홈 › Papers

Predicting Preferred Dialogue-to-Background Loudness Difference in Dialogue-Separated Audio

2023-05-30 · Luca Resti, Martin Strauss, Matteo Torcoli, Emanuël Habets, Bernd Edler

Dialogue Enhancement (DE) enables the rebalancing of dialogue and background sounds to fit personal preferences and needs in the context of broadcast audio. When individual audio stems are unavailable from production, Dialogue Separation (DS) can be applied to the final audio mixture to obtain estimates of these stems. This work focuses on Preferred Loudness Differences (PLDs) between dialogue and background sounds. While previous studies determined the PLD through a listening test employing original stems from production, stems estimated by DS are used in the present study. In addition, a larger variety of signal classes is considered. PLDs vary substantially across individuals (average interquartile range: 5.7 LU). Despite this variability, PLDs are found to be highly dependent on the signal type under consideration, and it is shown that median PLDs can be predicted using objective intelligibility metrics. Two existing baseline prediction methods - intended for use with original stems - displayed a Mean Absolute Error (MAE) of 7.5 LU and 5 LU, respectively. A modified baseline (MAE: 3.2 LU) and an alternative approach (MAE: 2.5 LU) are proposed. Results support the viability of processing final broadcast mixtures with DS and offering an alternative remixing that accounts for median PLDs.

📄 PDF Abstract BibTeX arXiv:2305.19100

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Dialogue Enhancement in Object-based Audio -- Evaluating the Benefit on People above 65

2020-06-25

Due to age-related hearing loss, elderly people often struggle with following the language on TV. Because they form an increasing part of the audience, this problem will become even more important in the future and needs…

blind source separation

Speech Loudness in Broadcasting and Streaming

2024-05-27 · Matteo Torcoli, Mhd Modar Halimeh, Thomas Leitz, Yannik Grewe 외

The introduction and regulation of loudness in broadcasting and streaming brought clear benefits to the audience, e.g., a level of uniformity across programs and channels. Yet, speech loudness is frequently reported as b…

Fast computation of loudness using a deep neural network

2019-05-24 · Josef Schlittenlacher, Richard E. Turner, Brian C. J. Moore

The present paper introduces a deep neural network (DNN) for predicting the instantaneous loudness of a sound from its time waveform. The DNN was trained using the output of a more complex model, called the Cambridge lou…

Vehicle Noise: Comparison of Loudness Ratings in the Field and the Laboratory

2022-04-28 · Gerard Llorach, Dirk Oetting, Matthias Vormann, Markus Meis 외

Objective: Distorted loudness perception is one of the main complaints of hearing aid users. Being able to measure loudness perception correctly in the clinic is essential for fitting hearing aids. For this, experiments …

Predicting Turn-Taking Outcomes in Multi-Party Conversation: Interpretable Modelling of Speech and Gaze Dynamics with Interpersonal Closeness

2026-08-28 · Mark Dourado, Karim Haddad, Henrik G. Hassager, Stefania Serafin arxiv

Smooth speaker transitions are fundamental to effective conversation and rely on an interlocutor's ability to predict when to enter the conversation. This ability depends on accurately interpreting and expressing the ver…