paper-with-me

Papers

A Time-Frequency Generative Adversarial based method for Audio Packet Loss Concealment

2023-07-28 · Carlo Aironi, Samuele Cornell, Luca Serafini, Stefano Squartini

Packet loss is a major cause of voice quality degradation in VoIP transmissions with serious impact on intelligibility and user experience. This paper describes a system based on a generative adversarial approach, which aims to repair the lost fragments during the transmission of audio streams. Inspired by the powerful image-to-image translation capability of Generative Adversarial Networks (GANs), we propose bin2bin, an improved pix2pix framework to achieve the translation task from magnitude spectrograms of audio frames with lost packets, to noncorrupted speech spectrograms. In order to better maintain the structural information after spectrogram translation, this paper introduces the combination of two STFT-based loss functions, mixed with the traditional GAN objective. Furthermore, we employ a modified PatchGAN structure as discriminator and we lower the concealment time by a proper initialization of the phase reconstruction algorithm. Experimental results show that the proposed method has obvious advantages when compared with the current state-of-the-art methods, as it can better handle both high packet loss rates and large gaps.

📄 PDF Abstract BibTeX arXiv:2307.15611

Code (1)

aircarlo/bin2bin-gan-plc 공식 구현 pytorch

Tasks

Image-to-Image TranslationPacket Loss ConcealmentTranslation

Methods 이 논문이 사용한 방법론

Repair 설명 없음
Batch Normalization 설명 없음
HuMan(Expedia)||How do I get a human at Expedia? How do I get a human at Expedia? How Do I Get a Human at Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Real-Time Help & Exclusive…
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Sigmoid Activation 설명 없음
Concatenated Skip Connection A Concatenated Skip Connection is a type of skip connection that seeks to reuse features by concatenating them to new layers, allowing more information to be retained from…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

Adversarial Generation of Time-Frequency Features with application in audio synthesis

2019-02-11 · 36th International Conference on Machine Learning 2019 6 · Andrés Marafioti, Nicki Holighaus, Nathanaël Perraudin, Piotr Majdak

Time-frequency (TF) representations provide powerful and intuitive features for the analysis of time series such as audio. But still, generative modeling of audio in the TF domain is a subtle matter. Consequently, neural…

Audio GenerationAudio SynthesisGenerative Adversarial NetworkTime Series+1

Synthetic Traffic Generation with Wasserstein Generative Adversarial Networks

2022-12-05 · IEEE Global Communications Conference 2022 12 · Chao–Lun Wu, Yu–Ying Chen, Po–Yu Chou, Chih–Yu Wang

Network traffic data are critical for network research. With the help of synthetic traffic, researchers can readily generate data for network simulation and performance evaluation. However, the state-of-the-art traffic g…

Intelligent CommunicationSynthetic Data Generation

Real-Time Packet Loss Concealment With Mixed Generative and Predictive Model

2022-05-11 · Jean-Marc Valin, Ahmed Mustafa, Christopher Montgomery, Timothy B. Terriberry 외

As deep speech enhancement algorithms have recently demonstrated capabilities greatly surpassing their traditional counterparts for suppressing noise, reverberation and echo, attention is turning to the problem of packet…

Packet Loss ConcealmentSpeech EnhancementSpeech Synthesis

Multi-level Wavelet-based Generative Adversarial Network for Perceptual Quality Enhancement of Compressed Video

2020-08-02 · ECCV 2020 8 · Jianyi Wang, Xin Deng, Mai Xu, Congyong Chen 외

The past few years have witnessed fast development in video quality enhancement via deep learning. Existing methods mainly focus on enhancing the objective quality of compressed video while ignoring its perceptual qualit…

Generative Adversarial NetworkMotion Compensation

On the Use of Audio Fingerprinting Features for Speech Enhancement with Generative Adversarial Network

2020-07-27 · Farnood Faraji, Yazid Attabi, Benoit Champagne, Wei-Ping Zhu

The advent of learning-based methods in speech enhancement has revived the need for robust and reliable training features that can compactly represent speech signals while preserving their vital information. Time-frequen…

Generative Adversarial NetworkSpeech Enhancement