paper-with-me

홈 › Papers

Unsupervised word segmentation and lexicon discovery using acoustic word embeddings

2016-03-09 · Herman Kamper, Aren Jansen, Sharon Goldwater

In settings where only unlabelled speech data is available, speech technology needs to be developed without transcriptions, pronunciation dictionaries, or language modelling text. A similar problem is faced when modelling infant language acquisition. In these cases, categorical linguistic structure needs to be discovered directly from speech audio. We present a novel unsupervised Bayesian model that segments unlabelled speech and clusters the segments into hypothesized word groupings. The result is a complete unsupervised tokenization of the input speech in terms of discovered word types. In our approach, a potential word segment (of arbitrary length) is embedded in a fixed-dimensional acoustic vector space. The model, implemented as a Gibbs sampler, then builds a whole-word acoustic model in this space while jointly performing segmentation. We report word error rates in a small-vocabulary connected digit recognition task by mapping the unsupervised decoded output to ground truth transcriptions. The model achieves around 20% error rate, outperforming a previous HMM-based system by about 10% absolute. Moreover, in contrast to the baseline, our model does not require a pre-specified vocabulary size.

📄 PDF Abstract BibTeX arXiv:1603.02845

Code (0)

등록된 구현이 없습니다.

Tasks

Language AcquisitionLanguage ModellingWord Embeddings

Similar Papers 제목 키워드 기반

Unsupervised Lexicon Discovery from Acoustic Input

2015-01-01 · TACL 2015 1 · Chia-Ying Lee, Timothy J. O{'}Donnell, James Glass

We present a model of unsupervised phonological lexicon discovery{---}the problem of simultaneously learning phoneme-like and word-like units from acoustic input. Our model builds on earlier models of unsupervised phone-…

Language AcquisitionSpeech Recognition

Revisiting speech segmentation and lexicon learning with better features

2024-01-31 · Herman Kamper, Benjamin van Niekerk

We revisit a self-supervised method that segments unlabelled speech into word-like segments. We start from the two-stage duration-penalised dynamic programming method that performs zero-resource segmentation without lear…

Acoustic Unit DiscoverySegmentation

Unsupervised Discovery of Linguistic Structure Including Two-level Acoustic Patterns Using Three Cascaded Stages of Iterative Optimization

2015-09-07 · Cheng-Tao Chung, Chun-an Chan, Lin-shan Lee

Techniques for unsupervised discovery of acoustic patterns are getting increasingly attractive, because huge quantities of speech data are becoming available but manual annotations remain hard to acquire. In this paper, …

Language ModelingLanguage Modelling

A segmental framework for fully-unsupervised large-vocabulary speech recognition

2016-06-22 · Herman Kamper, Aren Jansen, Sharon Goldwater

Zero-resource speech technology is a growing research area that aims to develop methods for speech processing in the absence of transcriptions, lexicons, or language modelling text. Early term discovery systems focused o…

Language ModellingSpeech RecognitionUnsupervised Speech Recognition

Bayesian Models for Unit Discovery on a Very Low Resource Language

2018-02-16 · Lucas Ondel, Pierre Godard, Laurent Besacier, Elin Larsen 외

Developing speech technologies for low-resource languages has become a very active research field over the last decade. Among others, Bayesian models have shown some promising results on artificial examples but still lac…

Acoustic Unit DiscoverySegmentation