paper-with-me

홈 › Papers

Improved Accent Classification Combining Phonetic Vowels with Acoustic Features

2016-02-24 · Zhenhao Ge

Researches have shown accent classification can be improved by integrating semantic information into pure acoustic approach. In this work, we combine phonetic knowledge, such as vowels, with enhanced acoustic features to build an improved accent classification system. The classifier is based on Gaussian Mixture Model-Universal Background Model (GMM-UBM), with normalized Perceptual Linear Predictive (PLP) features. The features are further optimized by Principle Component Analysis (PCA) and Hetroscedastic Linear Discriminant Analysis (HLDA). Using 7 major types of accented speech from the Foreign Accented English (FAE) corpus, the system achieves classification accuracy 54% with input test data as short as 20 seconds, which is competitive to the state of the art in this field.

📄 PDF Abstract BibTeX arXiv:1602.07394

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationGeneral Classification

Similar Papers 제목 키워드 기반

Accent Classification with Phonetic Vowel Representation

2016-02-24 · Zhenhao Ge, Yingyi Tan, Aravind Ganapathiraju

Previous accent classification research focused mainly on detecting accents with pure acoustic information without recognizing accented speech. This work combines phonetic knowledge such as vowels with acoustic informati…

ClassificationGeneral Classification

Multi-Accent Mandarin Dry-Vocal Singing Dataset: Benchmark for Singing Accent Recognition

2025-12-07 · Zihao Wang, Ruibin Yuan, Ziqi Geng, Hengjia Li 외 arxiv

Singing accent research is underexplored compared to speech accent studies, primarily due to the scarcity of suitable datasets. Existing singing datasets often suffer from detail loss, frequently resulting from the vocal…

On the Relationship between Accent Strength and Articulatory Features

2025-07-03 · Kevin Huang, Sean Foley, Jihwan Lee, Yoonjeong Lee 외 arxiv

This paper explores the relationship between accent strength and articulatory features inferred from acoustic speech. To quantify accent strength, we compare phonetic transcriptions with transcriptions based on dictionar…

Self-Supervised Learning

VOCAL: Vowel and Consonant Layering for Expressive Animator-Centric Singing Animation

2022-11-30 · Siggraph Asia 2022 2022 11 · Yifang Pan, Chris Landreth, Eugene Fiume, Karan Singh Authors Info & Claims

Singing and speaking are two fundamental forms of human communication. From a modeling perspective however, speaking can be seen as a subset of singing. We present VOCAL, a system that automatically generates expressive,…

Sensitivity

Accented Text-to-Speech Synthesis with Limited Data

2023-05-08 · Xuehao Zhou, Mingyang Zhang, Yi Zhou, Zhizheng Wu 외

This paper presents an accented text-to-speech (TTS) synthesis framework with limited training data. We study two aspects concerning accent rendering: phonetic (phoneme difference) and prosodic (pitch pattern and phoneme…

Speech Synthesistext-to-speechText to SpeechText-To-Speech Synthesis