paper-with-me

홈 › Papers

Enhanced Voice Post Processing Using Voice Decoder Guidance Indicators

2019-11-13

Voice enhancement and voice coding are imperative and important functions in a voice-communication system. However, both functions are commonly treated independently, even though both utilize similar features of the underlying signals. Our proposal is to leverage information from one function to the benefit of the other. Specifically, our proposed changes are focused on changes to the voice enhancement at the downlink side and utilizing information of the voice decoding. Preliminary results show that such an approach results in improved quality. Additionally, suggestions are provided on future extensions of the proposed concept.

📄 PDF Abstract BibTeX arXiv:1911.05560

Code (0)

등록된 구현이 없습니다.

Tasks

Decoder

Similar Papers 제목 키워드 기반

An Empirical Study on End-to-End Singing Voice Synthesis with Encoder-Decoder Architectures

2021-08-06 · Dengfeng Ke, Yuxing Lu, Xudong Liu, Yanyan Xu 외

With the rapid development of neural network architectures and speech processing models, singing voice synthesis with neural networks is becoming the cutting-edge technique of digital music production. In this work, in o…

DecoderSinging Voice Synthesis

Evaluation of Google's Voice Recognition and Sentence Classification for Health Care Applications

2024-02-02 · Majbah Uddin, Nathan Huynh, Jose M Vidal, Kevin M Taaffe 외

This study examined the use of voice recognition technology in perioperative services (Periop) to enable Periop staff to record workflow milestones using mobile technology. The use of mobile technology to improve patient…

SentenceSentence Classification

Real-Time and Accurate: Zero-shot High-Fidelity Singing Voice Conversion with Multi-Condition Flow Synthesis

2024-05-23 · Hui Li, Hongyu Wang, Zhijin Chen, Bohan Sun 외

Singing voice conversion is to convert the source singing voice into the target singing voice except for the content. Currently, flow-based models can complete the task of voice conversion, but they struggle to effective…

AttributeDecoderVoice Conversion

Voice Filter: Few-shot text-to-speech speaker adaptation using voice conversion as a post-processing module

2022-02-16 · Adam Gabryś, Goeric Huybrechts, Manuel Sam Ribeiro, Chung-Ming Chien 외

State-of-the-art text-to-speech (TTS) systems require several hours of recorded speech data to generate high-quality synthetic speech. When using reduced amounts of training data, standard TTS models suffer from speech q…

Speech Synthesistext-to-speechText to SpeechVoice Conversion

MinMo: A Multimodal Large Language Model for Seamless Voice Interaction

2025-01-10 · Qian Chen, Yafeng Chen, Yanni Chen, Mengzhe Chen 외

Recent advancements in large language models (LLMs) and multimodal speech-text models have laid the groundwork for seamless voice interactions, enabling real-time, natural, and human-like conversations. Previous models f…

Instruction FollowingLanguage ModelingLanguage ModellingLarge Language Model+4