paper-with-me

홈 › Papers

Low-rank and Sparse Soft Targets to Learn Better DNN Acoustic Models

2016-10-18 · Pranay Dighe, Afsaneh Asaei, Herve Bourlard

Conventional deep neural networks (DNN) for speech acoustic modeling rely on Gaussian mixture models (GMM) and hidden Markov model (HMM) to obtain binary class labels as the targets for DNN training. Subword classes in speech recognition systems correspond to context-dependent tied states or senones. The present work addresses some limitations of GMM-HMM senone alignments for DNN training. We hypothesize that the senone probabilities obtained from a DNN trained with binary labels can provide more accurate targets to learn better acoustic models. However, DNN outputs bear inaccuracies which are exhibited as high dimensional unstructured noise, whereas the informative components are structured and low-dimensional. We exploit principle component analysis (PCA) and sparse coding to characterize the senone subspaces. Enhanced probabilities obtained from low-rank and sparse reconstructions are used as soft-targets for DNN acoustic modeling, that also enables training with untranscribed data. Experiments conducted on AMI corpus shows 4.6% relative reduction in word error rate.

📄 PDF Abstract BibTeX arXiv:1610.05688

Code (0)

등록된 구현이 없습니다.

Tasks

speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Learning Better Structured Representations Using Low-rank Adaptive Label Smoothing

2021-01-01 · ICLR 2021 1 · Asish Ghoshal, Xilun Chen, Sonal Gupta, Luke Zettlemoyer 외

Training with soft targets instead of hard targets has been shown to improve performance and calibration of deep neural networks. Label smoothing is popular way of computing soft targets, where one-hot encoding of a clas…

Generalization BoundsMachine TranslationSemantic ParsingStructured Prediction+2

Scatterbrain: Unifying Sparse and Low-rank Attention Approximation

2021-10-28 · NeurIPS 2021 12 · Beidi Chen, Tri Dao, Eric Winsor, Zhao Song 외

Recent advances in efficient Transformers have exploited either the sparsity or low-rank properties of attention matrices to reduce the computational and memory bottlenecks of modeling long sequences. However, it is stil…

Image GenerationLanguage ModelingLanguage Modelling

Scatterbrain: Unifying Sparse and Low-rank Attention

2021-05-21 · NeurIPS 2021 12 · Beidi Chen, Tri Dao, Eric Winsor, Zhao Song 외

Recent advances in efficient Transformers have exploited either the sparsity or low-rank properties of attention matrices to reduce the computational and memory bottlenecks of modeling long sequences. However, it is stil…

Image GenerationLanguage ModelingLanguage Modelling

Sparse and Low-Rank Matrix Decomposition for Automatic Target Detection in Hyperspectral Imagery

2017-11-24 · Ahmad W. Bitar, Loong-Fah Cheong, Jean-Philippe Ovarlez

Given a target prior information, our goal is to propose a method for automatically separating targets of interests from the background in hyperspectral imagery. More precisely, we regard the given hyperspectral image (H…

SE-SSD: Self-Ensembling Single-Stage Object Detector From Point Cloud

2021-04-20 · CVPR 2021 1 · Wu Zheng, Weiliang Tang, Li Jiang, Chi-Wing Fu

We present Self-Ensembling Single-Stage object Detector (SE-SSD) for accurate and efficient 3D object detection in outdoor point clouds. Our key focus is on exploiting both soft and hard targets with our formulated const…

3D Object DetectionBirds Eye View Object DetectionObjectobject-detection+1