The "Sound of Silence" in EEG -- Cognitive voice activity detection
Speech cognition bears potential application as a brain computer interface that can improve the quality of life for the otherwise communication impaired people. While speech and resting state EEG are popularly studied, here we attempt to explore a "non-speech"(NS) state of brain activity corresponding to the silence regions of speech audio. Firstly, speech perception is studied to inspect the existence of such a state, followed by its identification in speech imagination. Analogous to how voice activity detection is employed to enhance the performance of speech recognition, the EEG state activity detection protocol implemented here is applied to boost the confidence of imagined speech EEG decoding. Classification of speech and NS state is done using two datasets collected from laboratory-based and commercial-based devices. The state sequential information thus obtained is further utilized to reduce the search space of imagined EEG unit recognition. Temporal signal structures and topographic maps of NS states are visualized across subjects and sessions. The recognition performance and the visual distinction observed demonstrates the existence of silence signatures in EEG.
Code (0)
등록된 구현이 없습니다.
Tasks
Action DetectionActivity DetectionBrain Computer InterfaceEEGEeg DecodingElectroencephalogram (EEG)speech-recognitionSpeech RecognitionSimilar Papers 제목 키워드 기반
Bts-e: Audio deepfake detection using breathing-talking-silence encoder
Voice phishing (vishing) is increasingly popular due to the development of speech synthesis technology. In particular, the use of deep learning to generate an arbitrary-content audio clip simulating the victim’s voice ma…
Audio Deepfake DetectionDeepFake DetectionFace SwappingSpeaker Verification+3The Impact of Silence on Speech Anti-Spoofing
The current speech anti-spoofing countermeasures (CMs) show excellent performance on specific datasets. However, removing the silence of test speech through Voice Activity Detection (VAD) can severely degrade performance…
Action DetectionActivity Detectiontext-to-speechText to Speech+1Property-Aware Multi-Speaker Data Simulation: A Probabilistic Modelling Technique for Synthetic Data Generation
We introduce a sophisticated multi-speaker speech data simulator, specifically engineered to generate multi-speaker speech recordings. A notable feature of this simulator is its capacity to modulate the distribution of s…
Action DetectionActivity Detectionspeaker-diarizationSpeaker Diarization+1Semantic VAD: Low-Latency Voice Activity Detection for Speech Interaction
For speech interaction, voice activity detection (VAD) is often used as a front-end. However, traditional VAD algorithms usually need to wait for a continuous tail silence to reach a preset maximum duration before segmen…
Action DetectionActivity DetectionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)+3How social feedback processing in the brain shapes collective opinion processes in the era of social media
What are the mechanisms by which groups with certain opinions gain public voice and force others holding a different view into silence? And how does social media play into this? Drawing on recent neuro-scientific insight…