paper-with-me

Papers

Bilevel Joint Unsupervised and Supervised Training for Automatic Speech Recognition

2024-12-11 · Xiaodong Cui, A F M Saif, Songtao Lu, Lisha Chen, Tianyi Chen, Brian Kingsbury, George Saon

In this paper, we propose a bilevel joint unsupervised and supervised training (BL-JUST) framework for automatic speech recognition. Compared to the conventional pre-training and fine-tuning strategy which is a disconnected two-stage process, BL-JUST tries to optimize an acoustic model such that it simultaneously minimizes both the unsupervised and supervised loss functions. Because BL-JUST seeks matched local optima of both loss functions, acoustic representations learned by the acoustic model strike a good balance between being generic and task-specific. We solve the BL-JUST problem using penalty-based bilevel gradient descent and evaluate the trained deep neural network acoustic models on various datasets with a variety of architectures and loss functions. We show that BL-JUST can outperform the widely-used pre-training and fine-tuning strategy and some other popular semi-supervised techniques.

📄 PDF Abstract BibTeX arXiv:2412.08548

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech Recognitionspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Joint Unsupervised and Supervised Training for Automatic Speech Recognition via Bilevel Optimization

2024-01-13 · A F M Saif, Xiaodong Cui, Han Shen, Songtao Lu 외

In this paper, we present a novel bilevel optimization-based training approach to training acoustic models for automatic speech recognition (ASR) tasks that we term {bi-level joint unsupervised and supervised training (B…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Bilevel Optimizationspeech-recognition+1

PANOM: Automatic Hyper-parameter Tuning for Inverse Problems

2021-10-19 · NeurIPS Workshop Deep_Invers 2021 12 · Tianci Liu, Quan Zhang, Qi Lei

Automated hyper-parameter tuning for unsupervised learning, including inverse problems, remains a long-standing open problem due to the lack of validation data. In this work, we design an automatic tuning criterion for i…

Bilevel Optimization

Bilevel Layer-Positioning LoRA for Real Image Dehazing

2026-03-11 · Yan Zhang, Long Ma, Yuxin Feng, Zhe Huang 외 arxiv

Learning-based real image dehazing methods have achieved notable progress, yet they still face adaptation challenges in diverse real haze scenes. These challenges mainly stem from the lack of effective unsupervised mecha…

Image Dehazing

Whiteness-based bilevel estimation of weighted TV parameter maps for image denoising

2025-03-10 · Monica Pragliola, Luca Calatroni, Alessandro Lanza

We consider a bilevel optimisation strategy based on normalised residual whiteness loss for estimating the weighted total variation parameter maps for denoising images corrupted by additive white Gaussian noise. Compared…

DenoisingImage Denoising

Learning Multi-level Sparse Representations

2013-12-01 · NeurIPS 2013 12 · Ferran Diego Andilla, Fred A. Hamprecht

Bilinear approximation of a matrix is a powerful paradigm of unsupervised learning. In some applications, however, there is a natural hierarchy of concepts that ought to be reflected in the unsupervised analysis. For exa…