paper-with-me

홈 › Papers

Phone Inventories and Recognition for Every Language

2022-06-01 · LREC 2022 6 · Xinjian Li, Florian Metze, David R. Mortensen, Alan W Black, Shinji Watanabe

Identifying phone inventories is a crucial component in language documentation and the preservation of endangered languages. However, even the largest collection of phone inventory only covers about 2000 languages, which is only 1/4 of the total number of languages in the world. A majority of the remaining languages are endangered. In this work, we attempt to solve this problem by estimating the phone inventory for any language listed in Glottolog, which contains phylogenetic information regarding 8000 languages. In particular, we propose one probabilistic model and one non-probabilistic model, both using phylogenetic trees (“language family trees”) to measure the distance between languages. We show that our best model outperforms baseline models by 6.5 F1. Furthermore, we demonstrate that, with the proposed inventories, the phone recognition model can be customized for every language in the set, which improved the PER (phone error rate) in phone recognition by 25%.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Allophant: Cross-lingual Phoneme Recognition with Articulatory Attributes

2023-06-07 · Kevin Glocker, Aaricia Herygers, Munir Georges

This paper proposes Allophant, a multilingual phoneme recognizer. It requires only a phoneme inventory for cross-lingual transfer to a target language, allowing for low-resource recognition. The architecture combines a c…

AttributeCross-Lingual TransferMulti-Task LearningPhoneme Recognition+2

Discovering Phonetic Inventories with Crosslingual Automatic Speech Recognition

2022-01-26 · Piotr Żelasko, Siyuan Feng, Laureano Moro Velazquez, Ali Abavisani 외

The high cost of data acquisition makes Automatic Speech Recognition (ASR) model training problematic for most existing languages, including languages that do not even have a written script, or for which the phone invent…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+2

Zero-shot Learning for Speech Recognition with Universal Phonetic Model

2018-09-27 · Xinjian Li, Siddharth Dalmia, David R. Mortensen, Florian Metze 외

There are more than 7,000 languages in the world, but due to the lack of training sets, only a small number of them have speech recognition systems. Multilingual speech recognition provides a solution if at least some au…

speech-recognitionSpeech RecognitionZero-Shot Learning

DiscoPhon: Benchmarking the Unsupervised Discovery of Phoneme Inventories With Discrete Speech Units

2026-03-19 · Maxime Poli, Manel Khentout, Angelo Ortiz Tandazo, Ewan Dunbar 외 arxiv

We introduce DiscoPhon, a multilingual benchmark for evaluating unsupervised phoneme discovery from discrete speech units. DiscoPhon covers 6 dev and 6 test languages, chosen to span a wide range of phonemic contrasts. G…

Universal Phone Recognition with a Multilingual Allophone System

2020-02-26 · Xinjian Li, Siddharth Dalmia, Juncheng Li, Matthew Lee 외

Multilingual models can improve language processing, particularly for low resource situations, by sharing parameters across languages. Multilingual acoustic models, however, generally ignore the difference between phonem…

speech-recognitionSpeech Recognition