paper-with-me

홈 › Papers

Selective Augmentation: Improving Universal Automatic Phonetic Transcription via G2P Bootstrapping

2026-04-29 · Tobias Bystrich, Julia M. Pritzen, Christoph A. Schmidt, Claudia Wich-Reif arxiv

In the field of universal automatic phonetic transcription (APT), clean and diverse training transcriptions are required. However, such high-quality data is limited. We propose the bootstrapping approach Selective Augmentation to improve the available training transcriptions by selectively transferring distinctions between languages. Based on the model MultIPA, we exemplarily show that we could increase the accuracy of an existing feature (plosive voicing) and add a new feature (plosive aspiration) by augmenting the existing training data using information from a separate helper language (Hindi). We describe intrinsic challenges of the evaluation and develop objective metrics to determine the success: Voicing accuracy was increased by 17.6% by reducing the number of false positives. Additionally, aspiration recognition was introduced: While the baseline transcribed 0% of German /p, t, k/ as aspirated, our approach transcribed them as aspirated in 61.2% of the cases. Introducing aspiration recognition to APT models allowed for the tenuis class to be successfully reduced by 32.2%, which also reduces the conflations between the test language's plosives.

📄 PDF Abstract BibTeX arXiv:2604.27204

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Universal Automatic Phonetic Transcription into the International Phonetic Alphabet

2023-08-07 · Chihiro Taguchi, Yusuke Sakai, Parisa Haghani, David Chiang

This paper presents a state-of-the-art model for transcribing speech in any language into the International Phonetic Alphabet (IPA). Transcription of spoken languages into IPA is an essential yet time-consuming process i…

A Universal System for Automatic Text-to-Phonetics Conversion

2019-09-01 · RANLP 2019 9 · Chen Gafni

This paper describes an automatic text-to-phonetics conversion system. The system was constructed to primarily serve as a research tool. It is implemented in a general-purpose linguistic software, which allows it to be i…

AlloVera: A Multilingual Allophone Database

2020-04-17 · LREC 2020 5 · David R. Mortensen, Xinjian Li, Patrick Littell, Alexis Michaud 외

We introduce a new resource, AlloVera, which provides mappings from 218 allophones to phonemes for 14 languages. Phonemes are contrastive phonological units, and allophones are their various concrete realizations, which …

speech-recognitionSpeech Recognition

Intent Recognition and Unsupervised Slot Identification for Low Resourced Spoken Dialog Systems

2021-04-03 · Akshat Gupta, Olivia Deng, Akruti Kushwaha, Saloni Mittal 외

Intent Recognition and Slot Identification are crucial components in spoken language understanding (SLU) systems. In this paper, we present a novel approach towards both these tasks in the context of low resourced and un…

Data AugmentationGeneral Classificationintent-classificationIntent Classification+3

Analysis of Phonetic Transcription for Danish Automatic Speech Recognition

2013-05-01 · WS 2013 5 · Andreas S{\o}eborg Kirkedal
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition