paper-with-me

Papers

Training Wake Word Detection with Synthesized Speech Data on Confusion Words

2020-11-03 · Yan Jia, Zexin Cai, Murong Ma, Zeqing Zhao, Xuyang Wang, Junjie Wang, Ming Li

Confusing-words are commonly encountered in real-life keyword spotting applications, which causes severe degradation of performance due to complex spoken terms and various kinds of words that sound similar to the predefined keywords. To enhance the wake word detection system's robustness on such scenarios, we investigate two data augmentation setups for training end-to-end KWS systems. One is involving the synthesized data from a multi-speaker speech synthesis system, and the other augmentation is performed by adding random noise to the acoustic feature. Experimental results show that augmentations help improve the system's robustness. Moreover, by augmenting the training set with the synthetic data generated by the multi-speaker text-to-speech system, we achieve a significant improvement regarding confusing words scenario.

📄 PDF Abstract BibTeX arXiv:2011.01460

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationKeyword SpottingSpeech Synthesistext-to-speechText to Speech

Similar Papers 제목 키워드 기반

Howl: A Deployed, Open-Source Wake Word Detection System

2020-08-21 · EMNLP (NLPOSS) 2020 11 · Raphael Tang, Jaejun Lee, Afsaneh Razi, Julia Cambre 외

We describe Howl, an open-source wake word detection toolkit with native support for open speech datasets, like Mozilla Common Voice and Google Speech Commands. We report benchmark results on Speech Commands and our own …

Keyword Spotting

Wake Word Detection with Alignment-Free Lattice-Free MMI

2020-05-17 · Yiming Wang, Hang Lv, Daniel Povey, Lei Xie 외

Always-on spoken language interfaces, e.g. personal digital assistants, rely on a wake word to start processing spoken input. We present novel methods to train a hybrid DNN/HMM wake word detection system from partially l…

Decoder

Low-resource Low-footprint Wake-word Detection using Knowledge Distillation

2022-07-06 · Arindam Ghosh, Mark Fuhs, Deblin Bagchi, Bahman Farahani 외

As virtual assistants have become more diverse and specialized, so has the demand for application or brand-specific wake words. However, the wake-word-specific datasets typically used to train wake-word detectors are cos…

Knowledge Distillationspeech-recognitionSpeech RecognitionTransfer Learning

Lightweight feature encoder for wake-up word detection based on self-supervised speech representation

2023-03-14 · Hyungjun Lim, Younggwan Kim, Kiho Yeom, Eunjoo Seo 외

Self-supervised learning method that provides generalized speech representations has recently received increasing attention. Wav2vec 2.0 is the most famous example, showing remarkable performance in numerous downstream s…

Dimensionality ReductionSelf-Supervised Learning

"OK Aura, Be Fair With Me": Demographics-Agnostic Training for Bias Mitigation in Wake-up Word Detection

2026-04-07 · Fernando López, Paula Delgado-Santos, Pablo Gómez, David Solans 외 arxiv

Voice-based interfaces are widely used; however, achieving fair Wake-up Word detection across diverse speaker populations remains a critical challenge due to persistent demographic biases. This study evaluates the effect…

Knowledge DistillationData Augmentation