paper-with-me

홈 › Papers

Dialogue Enhancement and Listening Effort in Broadcast Audio: A Multimodal Evaluation

2022-07-28 · Matteo Torcoli, Thomas Robotham, Emanuël A. P. Habets

Dialogue enhancement (DE) plays a vital role in broadcasting, enabling the personalization of the relative level between foreground speech and background music and effects. DE has been shown to improve the quality of experience, intelligibility, and self-reported listening effort (LE). A physiological indicator of LE known from audiology studies is pupil size. The relation between pupil size and LE is typically studied using artificial sentences and background noises not encountered in broadcast content. This work evaluates the effect of DE on LE in a multimodal manner that includes pupil size (tracked by a VR headset) and real-world audio excerpts from TV. Under ideal listening conditions, 28 normal-hearing participants listened to 30 audio excerpts presented in random order and processed by conditions varying the relative level between foreground and background audio. One of these conditions employed a recently proposed source separation system to attenuate the background given the original mixture as the sole input. After listening to each excerpt, subjects were asked to repeat the heard sentence and self-report the LE. Mean pupil dilation and peak pupil dilation were analyzed and compared with the self-report and the word recall rate. The multimodal evaluation shows a consistent trend of decreasing LE along with decreasing background level. DE, also when enabled by source separation, significantly reduces the pupil size as well as the self-reported LE. This highlights the benefit of personalization functionalities at the user's end.

📄 PDF Abstract BibTeX arXiv:2207.14240

Code (0)

등록된 구현이 없습니다.

Tasks

Pupil DilationSentence

Similar Papers 제목 키워드 기반

Reduction of Subjective Listening Effort for TV Broadcast Signals with Recurrent Neural Networks

2021-11-02 · Nils L. Westhausen, Rainer Huber, Hannah Baumgartner, Ragini Sinha 외

Listening to the audio of TV broadcast signals can be challenging for hearing-impaired as well as normal-hearing listeners, especially when background sounds are prominent or too loud compared to the speech signal. This …

Audio Source SeparationSpeech Enhancement

Predicting Preferred Dialogue-to-Background Loudness Difference in Dialogue-Separated Audio

2023-05-30 · Luca Resti, Martin Strauss, Matteo Torcoli, Emanuël Habets 외

Dialogue Enhancement (DE) enables the rebalancing of dialogue and background sounds to fit personal preferences and needs in the context of broadcast audio. When individual audio stems are unavailable from production, Di…

Dialog+ in Broadcasting: First Field Tests Using Deep-Learning-Based Dialogue Enhancement

2021-12-17 · Matteo Torcoli, Christian Simon, Jouni Paulus, Davide Straninger 외

Difficulties in following speech due to loud background sounds are common in broadcasting. Object-based audio, e.g., MPEG-H Audio solves this problem by providing a user-adjustable speech level. While object-based audio …

Object

HEAR: Hearing Enhanced Audio Response for Video-grounded Dialogue

2023-12-15 · Sunjae Yoon, Dahyun Kim, Eunseop Yoon, Hee Suk Yoon 외

Video-grounded Dialogue (VGD) aims to answer questions regarding a given multi-modal input comprising video, audio, and dialogue history. Although there have been numerous efforts in developing VGD systems to improve the…

Cooperative Audio Source Separation and Enhancement Using Distributed Microphone Arrays and Wearable Devices

2019-12-10

Augmented listening devices such as hearing aids often perform poorly in noisy and reverberant environments with many competing sound sources. Large distributed microphone arrays can improve performance, but data from re…

Audio Source Separation