paper-with-me

홈 › Papers

Truly unsupervised acoustic word embeddings using weak top-down constraints in encoder-decoder models

2018-11-01 · Herman Kamper

We investigate unsupervised models that can map a variable-duration speech segment to a fixed-dimensional representation. In settings where unlabelled speech is the only available resource, such acoustic word embeddings can form the basis for "zero-resource" speech search, discovery and indexing systems. Most existing unsupervised embedding methods still use some supervision, such as word or phoneme boundaries. Here we propose the encoder-decoder correspondence autoencoder (EncDec-CAE), which, instead of true word segments, uses automatically discovered segments: an unsupervised term discovery system finds pairs of words of the same unknown type, and the EncDec-CAE is trained to reconstruct one word given the other as input. We compare it to a standard encoder-decoder autoencoder (AE), a variational AE with a prior over its latent embedding, and downsampling. EncDec-CAE outperforms its closest competitor by 24% relative in average precision on two languages in a word discrimination task.

📄 PDF Abstract BibTeX arXiv:1811.00403

Code (2)

kamperh/recipe_bucktsong_awe tf
kamperh/recipe_bucktsong_awe_py3 tf

Tasks

DecoderWord Embeddings

Methods 이 논문이 사용한 방법론

AE An autoencoder is a type of artificial neural network used to learn efficient data codings in an unsupervised manner. The aim of an autoencoder is to learn a representation…
Solana Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

A comparison of self-supervised speech representations as input features for unsupervised acoustic word embeddings

2020-12-14 · Lisa van Staden, Herman Kamper

Many speech processing tasks involve measuring the acoustic similarity between speech segments. Acoustic word embeddings (AWE) allow for efficient comparisons by mapping speech segments of arbitrary duration to fixed-dim…

Representation LearningWord Embeddings

A Correspondence Variational Autoencoder for Unsupervised Acoustic Word Embeddings

2020-12-03 · Puyuan Peng, Herman Kamper, Karen Livescu

We propose a new unsupervised model for mapping a variable-duration speech segment to a fixed-dimensional representation. The resulting acoustic word embeddings can form the basis of search, discovery, and indexing syste…

Word Embeddings

Unsupervised Cross-Lingual Representation Learning

2019-07-01 · ACL 2019 7 · Sebastian Ruder, Anders S{\o}gaard, Ivan Vuli{\'c}

In this tutorial, we provide a comprehensive survey of the exciting recent work on cutting-edge weakly-supervised and unsupervised cross-lingual word representations. After providing a brief history of supervised cross-l…

Representation LearningStructured Prediction

Unsupervised word segmentation and lexicon discovery using acoustic word embeddings

2016-03-09 · Herman Kamper, Aren Jansen, Sharon Goldwater

In settings where only unlabelled speech data is available, speech technology needs to be developed without transcriptions, pronunciation dictionaries, or language modelling text. A similar problem is faced when modellin…

Language AcquisitionLanguage ModellingWord Embeddings

Unsupervised Cross-Lingual Part-of-Speech Tagging for Truly Low-Resource Scenarios

2020-11-01 · EMNLP 2020 11 · Ramy Eskander, Smaranda Muresan, Michael Collins

We describe a fully unsupervised cross-lingual transfer approach for part-of-speech (POS) tagging under a truly low resource scenario. We assume access to parallel translations between the target language and one or more…

Cross-Lingual TransferPart-Of-Speech TaggingPOSPOS Tagging+2