paper-with-me

Papers

Topology combined machine learning for consonant recognition

2023-11-26 · Pingyao Feng, Siheng Yi, Qingrui Qu, Zhiwang Yu, Yifei Zhu

In artificial-intelligence-aided signal processing, existing deep learning models often exhibit a black-box structure, and their validity and comprehensibility remain elusive. The integration of topological methods, despite its relatively nascent application, serves a dual purpose of making models more interpretable as well as extracting structural information from time-dependent data for smarter learning. Here, we provide a transparent and broadly applicable methodology, TopCap, to capture the most salient topological features inherent in time series for machine learning. Rooted in high-dimensional ambient spaces, TopCap is capable of capturing features rarely detected in datasets with low intrinsic dimensionality. Applying time-delay embedding and persistent homology, we obtain descriptors which encapsulate information such as the vibration of a time series, in terms of its variability of frequency, amplitude, and average line, demonstrated with simulated data. This information is then vectorised and fed into multiple machine learning algorithms such as k-nearest neighbours and support vector machine. Notably, in classifying voiced and voiceless consonants, TopCap achieves an accuracy exceeding 96% and is geared towards designing topological convolutional layers for deep learning of speech and audio signals.

📄 PDF Abstract BibTeX arXiv:2311.15210

Code (0)

등록된 구현이 없습니다.

Tasks

Time Series

Similar Papers 제목 키워드 기반

Persian Vowel recognition with MFCC and ANN on PCVC speech dataset

2018-12-17 · Saber Malekzadeh, Mohammad Hossein Gholizadeh, Seyed Naser Razavi

In this paper a new method for recognition of consonant-vowel phonemes combination on a new Persian speech dataset titled as PCVC (Persian Consonant-Vowel Combination) is proposed which is used to recognize Persian phone…

Phoneme Recognition

IR-UWB Radar-Based Contactless Silent Speech Recognition of Vowels, Consonants, Words, and Phrases

2023-12-15 · Sunghwa Lee, Younghoon Shin, Myungjong Kim, Jiwon Seo

Several sensing techniques have been proposed for silent speech recognition (SSR); however, many of these methods require invasive processes or sensor attachment to the skin using adhesive tape or glue, rendering them un…

Dynamic Time WarpingSilent Speech Recognitionspeech-recognitionSpeech Recognition

La reconnaissance des sons consonantiques en cas de d\'esynchronisation spectrale : avec et sans information spectrale fine (Recognition of desynchronized consonantics sounds with and without fine spectral structure) [in French]

2012-06-01 · JEPTALNRECITAL 2012 6 · Marjolaine Ray, Olivier Crouzet

Automatic Estimation of Intelligibility Measure for Consonants in Speech

2020-05-12 · Ali Abavisani, Mark Hasegawa-Johnson

In this article, we provide a model to estimate a real-valued measure of the intelligibility of individual speech segments. We trained regression models based on Convolutional Neural Networks (CNN) for stop consonants \t…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Pretrained self-supervised speech models can recognize unseen consonants

2026-06-10 · Chihiro Taguchi, Éric Le Ferrand, Hirosi Nakagawa, Hitomi Ono 외 arxiv

Modern pretrained self-supervised automatic speech recognition models are trained on large-scale audio data to encode speech into contextualized representations. However, their training data are heavily skewed toward hig…

Speech Recognition