paper-with-me

홈 › Papers

Learned Transferable Architectures Can Surpass Hand-Designed Architectures for Large Scale Speech Recognition

2020-08-25 · Liqiang He, Dan Su, Dong Yu

In this paper, we explore the neural architecture search (NAS) for automatic speech recognition (ASR) systems. With reference to the previous works in the computer vision field, the transferability of the searched architecture is the main focus of our work. The architecture search is conducted on the small proxy dataset, and then the evaluation network, constructed with the searched architecture, is evaluated on the large dataset. Especially, we propose a revised search space for speech recognition tasks which theoretically facilitates the search algorithm to explore the architectures with low complexity. Extensive experiments show that: (i) the architecture searched on the small proxy dataset can be transferred to the large dataset for the speech recognition tasks. (ii) the architecture learned in the revised search space can greatly reduce the computational overhead and GPU memory usage with mild performance degradation. (iii) the searched architecture can achieve more than 20% and 15% (average on the four test sets) relative improvements respectively on the AISHELL-2 dataset and the large (10k hours) dataset, compared with our best hand-designed DFSMN-SAN architecture. To the best of our knowledge, this is the first report of NAS results with large scale dataset (up to 10K hours), indicating the promising application of NAS to industrial ASR systems.

📄 PDF Abstract BibTeX arXiv:2008.11589

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)GPUNeural Architecture Searchspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

MLR-SNet: Transferable LR Schedules for Heterogeneous Tasks

2020-07-29 · Jun Shu, Yanwen Zhu, Qian Zhao, Zongben Xu 외

The learning rate (LR) is one of the most important hyper-parameters in stochastic gradient descent (SGD) algorithm for training deep neural networks (DNN). However, current hand-designed LR schedules need to manually pr…

text-classificationText Classification

LiDAR: Sensing Linear Probing Performance in Joint Embedding SSL Architectures

2023-12-07 · Vimal Thilak, Chen Huang, Omid Saremi, Laurent Dinh 외

Joint embedding (JE) architectures have emerged as a promising avenue for acquiring transferable data representations. A key obstacle to using JE methods, however, is the inherent challenge of evaluating learned represen…

TMLC-Net: Transferable Meta Label Correction for Noisy Label Learning

2025-02-11 · Mengyang Li

The prevalence of noisy labels in real-world datasets poses a significant impediment to the effective deployment of deep learning models. While meta-learning strategies have emerged as a promising approach for addressing…

Meta-Learning

HUB: Guiding Learned Optimizers with Continuous Prompt Tuning

2023-05-26 · Gaole Dai, Wei Wu, Ziyu Wang, Jie Fu 외

Learned optimizers are a crucial component of meta-learning. Recent advancements in scalable learned optimizers have demonstrated their superior performance over hand-designed optimizers in various tasks. However, certai…

Meta-Learning

Learning Transferable Architectures for Scalable Image Recognition

2017-07-21 · CVPR 2018 6 · Barret Zoph, Vijay Vasudevan, Jonathon Shlens, Quoc V. Le

Developing neural network image classification models often requires significant architecture engineering. In this paper, we study a method to learn the model architectures directly on the dataset of interest. As this ap…

Classificationimage-classificationImage ClassificationNeural Architecture Search