Zero-Bit Transmission of Adaptive Pre- and De-emphasis Filters for Speech and Audio Coding
This paper introduces a novel adaptation approach for first-order pre- and de-emphasis filters, an essential tool in many speech and audio codecs to increase coding efficiency and perceived quality. The proposed zero-bit self-adaptation approach differs from classical forward and backward adaptation approaches in that the de-emphasis coefficient is estimated at the receiver, from the decoded pre-emphasized signal. This eliminates the need to transmit information that arises from forward adaptation as well as the signal-filter lag that is inherent in backward adaptation. Evaluation results show that the de-emphasis coefficient can be estimated accurately from the decoded pre-emphasized signal and that the proposed zero-bit self-adaptation approach provides comparable subjective improvement to forward adaptation.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis
A good co-speech motion generation cannot be achieved without a careful integration of common rhythmic motion and rare yet essential semantic motion. In this work, we propose SemTalk for holistic co-speech motion generat…
Gesture GenerationMotion GenerationRhythmWideband Bandpass Filters Using a Novel Thick Metallization Technology
A new class of wideband bandpass filters based on using thick metallic bars as microwave resonators, instead of common microstrip lines, is presented. These bars provide a series of advantages over fully planar printed t…
LACE: A light-weight, causal model for enhancing coded speech through adaptive convolutions
Classical speech coding uses low-complexity postfilters with zero lookahead to enhance the quality of coded speech, but their effectiveness is limited by their simplicity. Deep Neural Networks (DNNs) can be much more eff…
Controllable Emphasis with zero data for text-to-speech
We present a scalable method to produce high quality emphasis for text-to-speech (TTS) that does not require recordings or annotations. Many TTS models include a phoneme duration model. A simple but effective method to a…
Sentencetext-to-speechText to SpeechCombinations of Adaptive Filters
Adaptive filters are at the core of many signal processing applications, ranging from acoustic noise supression to echo cancelation, array beamforming, channel equalization, to more recent sensor network applications in …
Mixture-of-Experts