paper-with-me

Papers

The Deterministic plus Stochastic Model of the Residual Signal and its Applications

2019-12-29 · Thomas Drugman, Thierry Dutoit

The modeling of speech production often relies on a source-filter approach. Although methods parameterizing the filter have nowadays reached a certain maturity, there is still a lot to be gained for several speech processing applications in finding an appropriate excitation model. This manuscript presents a Deterministic plus Stochastic Model (DSM) of the residual signal. The DSM consists of two contributions acting in two distinct spectral bands delimited by a maximum voiced frequency. Both components are extracted from an analysis performed on a speaker-dependent dataset of pitch-synchronous residual frames. The deterministic part models the low-frequency contents and arises from an orthonormal decomposition of these frames. As for the stochastic component, it is a high-frequency noise modulated both in time and frequency. Some interesting phonetic and computational properties of the DSM are also highlighted. The applicability of the DSM in two fields of speech processing is then studied. First, it is shown that incorporating the DSM vocoder in HMM-based speech synthesis enhances the delivered quality. The proposed approach turns out to significantly outperform the traditional pulse excitation and provides a quality equivalent to STRAIGHT. In a second application, the potential of glottal signatures derived from the proposed DSM is investigated for speaker identification purpose. Interestingly, these signatures are shown to lead to better recognition rates than other glottal-based methods.

📄 PDF Abstract BibTeX arXiv:2001.01000

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker IdentificationSpeech Synthesis

Similar Papers 제목 키워드 기반

A Deterministic plus Stochastic Model of the Residual Signal for Improved Parametric Speech Synthesis

2019-12-29 · Thomas Drugman, Geoffrey Wilfart, Thierry Dutoit

Speech generated by parametric synthesizers generally suffers from a typical buzziness, similar to what was encountered in old LPC-like vocoders. In order to alleviate this problem, a more suited modeling of the excitati…

Speech Synthesis

A Comparative Evaluation of Pitch Modification Techniques

2020-01-02 · Thomas Drugman, Thierry Dutoit

This paper addresses the problem of pitch modification, as an important module for an efficient voice transformation system. The Deterministic plus Stochastic Model of the residual signal we proposed in a previous work i…

NeuralDPS: Neural Deterministic Plus Stochastic Model with Multiband Excitation for Noise-Controllable Waveform Generation

2022-03-05 · Tao Wang, Ruibo Fu, Jiangyan Yi, JianHua Tao 외

The traditional vocoders have the advantages of high synthesis efficiency, strong interpretability, and speech editability, while the neural vocoders have the advantage of high synthesis quality. To combine the advantage…

CPU

Geometric bias in eigenspace perturbation under random heterogeneous noise

2026-06-09 · Fengkai Liu, Ke Wang, Wanjie Wang arxiv

Spectral methods rely fundamentally on the stability of principal eigenspaces under random perturbations. Classically, this stability is quantified by the Davis-Kahan and Wedin theorems, which bound the eigenspace error …

Use of operator defect identities in multi-channel signal plus residual-analysis via iterated products and telescoping energy-residuals: Applications to kernels in machine learning

2026-01-26 · Palle E. T. Jorgensen, Myung-Sin Song, James F. Tian arxiv

We present a new operator theoretic framework for analysis of complex systems with intrinsic subdivisions into components, taking the form of "residuals" in general, and "telescoping energy residuals" in particular. We p…