paper-with-me

홈 › Papers

Human and Machine Speaker Recognition Based on Short Trivial Events

2017-11-15 · Miao Zhang, Xiaofei Kang, Yanqing Wang, Lantian Li, Zhiyuan Tang, Haisheng Dai, Dong Wang

Trivial events are ubiquitous in human to human conversations, e.g., cough, laugh and sniff. Compared to regular speech, these trivial events are usually short and unclear, thus generally regarded as not speaker discriminative and so are largely ignored by present speaker recognition research. However, these trivial events are highly valuable in some particular circumstances such as forensic examination, as they are less subjected to intentional change, so can be used to discover the genuine speaker from disguised speech. In this paper, we collect a trivial event speech database that involves 75 speakers and 6 types of events, and report preliminary speaker recognition results on this database, by both human listeners and machines. Particularly, the deep feature learning technique recently proposed by our group is utilized to analyze and recognize the trivial events, which leads to acceptable equal error rates (EERs) despite the extremely short durations (0.2-0.5 seconds) of these events. Comparing different types of events, 'hmm' seems more speaker discriminative.

📄 PDF Abstract BibTeX arXiv:1711.05443

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Recognition

Similar Papers 제목 키워드 기반

Speaker Recognition with Cough, Laugh and "Wei"

2017-06-22 · Miao Zhang, Yixiang Chen, Lantian Li, Dong Wang

This paper proposes a speaker recognition (SRE) task with trivial speech events, such as cough and laugh. These trivial events are ubiquitous in conversations and less subjected to intentional change, therefore offering …

Speaker Recognition

"Hello, It's Me": Deep Learning-based Speech Synthesis Attacks in the Real World

2021-09-20 · Emily Wenger, Max Bronckers, Christian Cianfarani, Jenna Cryan 외

Advances in deep learning have introduced a new wave of voice synthesis tools, capable of producing audio that sounds as if spoken by a target speaker. If successful, such tools in the wrong hands will enable a range of …

Deep LearningSpeaker RecognitionSpeech Synthesis

Speaker Emotion Recognition: Leveraging Self-Supervised Models for Feature Extraction Using Wav2Vec2 and HuBERT

2024-11-05 · Pourya Jafarzadeh, Amir Mohammad Rostami, Padideh Choobdar

Speech is the most natural way of expressing ourselves as humans. Identifying emotion from speech is a nontrivial task due to the ambiguous definition of emotion itself. Speaker Emotion Recognition (SER) is essential for…

Emotion Recognition

A Deep Neural Network for Short-Segment Speaker Recognition

2019-07-22 · Amirhossein Hajavi, Ali Etemad

Todays interactive devices such as smart-phone assistants and smart speakers often deal with short-duration speech segments. As a result, speaker recognition systems integrated into such devices will be much better suite…

Speaker Recognition

Asynchronous Voice Anonymization Using Adversarial Perturbation On Speaker Embedding

2024-06-12 · Rui Wang, Liping Chen, Kong Aik Lee, Zhen-Hua Ling

Voice anonymization has been developed as a technique for preserving privacy by replacing the speaker's voice in a speech signal with that of a pseudo-speaker, thereby obscuring the original voice attributes from machine…

Disentanglement