paper-with-me

Papers

Unsupervised Adaptation with Interpretable Disentangled Representations for Distant Conversational Speech Recognition

2018-06-13 · Wei-Ning Hsu, Hao Tang, James Glass

The current trend in automatic speech recognition is to leverage large amounts of labeled data to train supervised neural network models. Unfortunately, obtaining data for a wide range of domains to train robust models can be costly. However, it is relatively inexpensive to collect large amounts of unlabeled data from domains that we want the models to generalize to. In this paper, we propose a novel unsupervised adaptation method that learns to synthesize labeled data for the target domain from unlabeled in-domain data and labeled out-of-domain data. We first learn without supervision an interpretable latent representation of speech that encodes linguistic and nuisance factors (e.g., speaker and channel) using different latent variables. To transform a labeled out-of-domain utterance without altering its transcript, we transform the latent nuisance variables while maintaining the linguistic variables. To demonstrate our approach, we focus on a channel mismatch setting, where the domain of interest is distant conversational speech, and labels are only available for close-talking speech. Our proposed method is evaluated on the AMI dataset, outperforming all baselines and bridging the gap between unadapted and in-domain models by over 77% without using any parallel data.

📄 PDF Abstract BibTeX arXiv:1806.04872

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Inferencing Based on Unsupervised Learning of Disentangled Representations

2018-03-07 · Tobias Hinz, Stefan Wermter

Combining Generative Adversarial Networks (GANs) with encoders that learn to encode data points has shown promising results in learning data representations in an unsupervised way. We propose a framework that combines an…

DescriptiveRepresentation LearningUnsupervised Image ClassificationUnsupervised MNIST

Disentangled and Self-Explainable Node Representation Learning

2024-10-28 · Simone Piaggesi, André Panisson, Megha Khosla

Node representations, or embeddings, are low-dimensional vectors that capture node properties, typically learned through unsupervised structural similarity objectives or supervised tasks. While recent efforts have focuse…

DisentanglementRepresentation Learning

Where and What? Examining Interpretable Disentangled Representations

2021-04-07 · CVPR 2021 1 · Xinqi Zhu, Chang Xu, DaCheng Tao

Capturing interpretable variations has long been one of the goals in disentanglement learning. However, unlike the independence assumption, interpretability has rarely been exploited to encourage disentanglement in the u…

DisentanglementModel SelectionPerceptual Distance

The Context-Aware Learner

2018-01-01 · ICLR 2018 1 · Conor Durkan, Amos Storkey, Harrison Edwards

One important aspect of generalization in machine learning involves reasoning about previously seen data in new settings. Such reasoning requires learning disentangled representations of data which are interpretable in i…

Meta-Learning

Learning Interpretable and Discrete Representations with Adversarial Training for Unsupervised Text Classification

2020-04-28 · Yau-Shian Wang, Hung-Yi Lee, Yun-Nung Chen

Learning continuous representations from unlabeled textual data has been increasingly studied for benefiting semi-supervised learning. Although it is relatively easier to interpret discrete representations, due to the di…

General Classificationtext-classificationText ClassificationUnsupervised Text Classification