paper-with-me

홈 › Papers

Psychoacoustic Challenges Of Speech Enhancement On VoIP Platforms

2023-10-11 · Joseph Konan, Shikhar Agnihotri, Ojas Bhargave, Shuo Han, Yunyang Zeng, Ankit Shah, Bhiksha Raj

Within the ambit of VoIP (Voice over Internet Protocol) telecommunications, the complexities introduced by acoustic transformations merit rigorous analysis. This research, rooted in the exploration of proprietary sender-side denoising effects, meticulously evaluates platforms such as Google Meets and Zoom. The study draws upon the Deep Noise Suppression (DNS) 2020 dataset, ensuring a structured examination tailored to various denoising settings and receiver interfaces. A methodological novelty is introduced via Blinder-Oaxaca decomposition, traditionally an econometric tool, repurposed herein to analyze acoustic-phonetic perturbations within VoIP systems. To further ground the implications of these transformations, psychoacoustic metrics, specifically PESQ and STOI, were used to explain of perceptual quality and intelligibility. Cumulatively, the insights garnered underscore the intricate landscape of VoIP-influenced acoustic dynamics. In addition to the primary findings, a multitude of metrics are reported, extending the research purview. Moreover, out-of-domain benchmarking for both time and time-frequency domain speech enhancement models is included, thereby enhancing the depth and applicability of this inquiry.

📄 PDF Abstract BibTeX arXiv:2310.07161

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingDenoisingSpeech Enhancement

Similar Papers 제목 키워드 기반

Improving Perceptual Quality, Intelligibility, and Acoustics on VoIP Platforms

2023-03-16 · Joseph Konan, Ojas Bhargave, Shikhar Agnihotri, Hojeong Lee 외

In this paper, we present a method for fine-tuning models trained on the Deep Noise Suppression (DNS) 2020 Challenge to improve their performance on Voice over Internet Protocol (VoIP) applications. Our approach involves…

Multi-Task LearningSpeech Enhancementspeech-recognitionSpeech Recognition

Cellular Network Speech Enhancement: Removing Background and Transmission Noise

2023-01-22 · Amanda Shu, Hamza Khalid, Haohui Liu, Shikhar Agnihotri 외

The primary objective of speech enhancement is to reduce background noise while preserving the target's speech. A common dilemma occurs when a speaker is confined to a noisy environment and receives a call with high back…

Speech Enhancement

DeepFilterNet: Perceptually Motivated Real-Time Speech Enhancement

2023-05-14 · Hendrik Schröter, Tobias Rosenkranz, Alberto N. Escalante-B., Andreas Maier

Multi-frame algorithms for single-channel speech enhancement are able to take advantage from short-time correlations within the speech signal. Deep Filtering (DF) was proposed to directly estimate a complex filter in fre…

CPUSpeech Enhancement

Are Modern Speech Enhancement Systems Vulnerable to Adversarial Attacks?

2025-09-25 · Rostislav Makarov, Lea Schönherr, Timo Gerkmann arxiv

Machine learning approaches for speech enhancement are becoming increasingly expressive, enabling ever more powerful modifications of input signals. In this paper, we demonstrate that this expressiveness introduces a vul…

Speech Enhancement

Real-Time Steganalysis for Stream Media Based on Multi-channel Convolutional Sliding Windows

2019-02-04 · Zhongliang Yang, Hao Yang, Yuting Hu, Yongfeng Huang 외

Previous VoIP steganalysis methods face great challenges in detecting speech signals at low embedding rates, and they are also generally difficult to perform real-time detection, making them hard to truly maintain cybers…

Steganalysis