paper-with-me

홈 › Papers

Phoneme Alignment Using the Information on Phonological Processes in Continuous Speech

2016-05-01 · LREC 2016 5 · Daniil Kocharov

The current study focuses on optimization of Levenshtein algorithm for the purpose of computing the optimal alignment between two phoneme transcriptions of spoken utterance containing sequences of phonetic symbols. The alignment is computed with the help of a confusion matrix in which costs for phonetic symbol deletion, insertion and substitution are defined taking into account various phonological processes that occur in fluent speech, such as anticipatory assimilation, phone elision and epenthesis. The corpus containing about 30 hours of Russian read speech was used to evaluate the presented algorithms. The experimental results have shown significant reduction of misalignment rate in comparison with the baseline Levenshtein algorithm: the number of errors has been reduced from 1.1 {\%} to 0.28 {\%}

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Phoneme Similarity Matrices to Improve Long Audio Alignment for Automatic Subtitling

2014-05-01 · LREC 2014 5 · Pablo Ruiz, Aitor {\'A}lvarez, Haritz Arzelus

Long audio alignment systems for Spanish and English are presented, within an automatic subtitling application. Language-specific phone decoders automatically recognize audio contents at phoneme level. At the same time, …

DecoderLanguage Modelling

Multilingual Dysarthric Speech Assessment Using Universal Phone Recognition and Language-Specific Phonemic Contrast Modeling

2026-01-29 · Eunjung Yeo, Julie M. Liss, Visar Berisha, David R. Mortensen arxiv

The growing prevalence of neurological disorders associated with dysarthria motivates the need for automated intelligibility assessment methods that are applicalbe across languages. However, most existing approaches are …

Learning-free L2-Accented Speech Generation using Phonological Rules

2026-03-08 · Thanathai Lertpetchpun, Yoonjeong Lee, Jihwan Lee, Tiantian Feng 외 arxiv

Accent plays a crucial role in speaker identity and inclusivity in speech technologies. Existing accented text-to-speech (TTS) systems either require large-scale accented datasets or lack fine-grained phoneme-level contr…

Modelling the Diachronic Emergence of Phoneme Frequency Distributions

2026-03-10 · Fermín Moscoso del Prado Martín, Suchir Salhan arxiv

Phoneme frequency distributions exhibit robust statistical regularities across languages, including exponential-tailed rank-frequency patterns and a negative relationship between phonemic inventory size and the relative …

Morphological Inflection with Phonological Features

2023-06-21 · David Guriel, Omer Goldman, Reut Tsarfaty

Recent years have brought great advances into solving morphological tasks, mostly due to powerful neural models applied to various tasks as (re)inflection and analysis. Yet, such morphological tasks cannot be considered …

Morphological Inflection