Toward Low-Cost End-to-End Spoken Language Understanding
Recent advances in spoken language understanding benefited from Self-Supervised models trained on large speech corpora. For French, the LeBenchmark project has made such models available and has led to impressive progress on several tasks including spoken language understanding. These advances have a non-negligible cost in terms of computation time and energy consumption. In this paper, we compare several learning strategies trying to reduce such cost while keeping competitive performance. At the same time we propose an extensive analysis where we measure the cost of our models in terms of training time and electric energy consumption, hopefully promoting a comprehensive evaluation procedure. The experiments are performed on the FSC and MEDIA corpora, and show that it is possible to reduce the learning cost while maintaining state-of-the-art performance and using SSL models.
Code (0)
등록된 구현이 없습니다.
Tasks
Spoken Language UnderstandingSimilar Papers 제목 키워드 기반
Vers la compréhension automatique de la parole bout-en-bout à moindre effort
Recent advances in spoken language understanding benefited from Self-Supervised models trained on large speech corpora. For French, the LeBenchmark project has made such models available and has led to impressive progres…
Spoken Language UnderstandingEnd-to-End Spoken Language Understanding Without Full Transcripts
An essential component of spoken language understanding (SLU) is slot filling: representing the meaning of a spoken utterance using semantic entity labels. In this paper, we develop end-to-end (E2E) spoken language under…
Decoderslot-fillingSlot Fillingspeech-recognition+2Cross-lingual transfer learning for spoken language understanding
Typically, spoken language understanding (SLU) models are trained on annotated data which are costly to gather. Aiming to reduce data needs for bootstrapping a SLU system for a new language, we present a simple but effec…
Cross-Lingual TransferSpoken Language UnderstandingTransfer LearningAugmenting Slot Values and Contexts for Spoken Language Understanding with Pretrained Models
Spoken Language Understanding (SLU) is one essential step in building a dialogue system. Due to the expensive cost of obtaining the labeled data, SLU suffers from the data scarcity problem. Therefore, in this paper, we f…
Data Augmentationslot-fillingSlot FillingSpoken Language UnderstandingLeveraging study of robustness and portability of spoken language understanding systems across languages and domains: the PORTMEDIA corpora
The PORTMEDIA project is intended to develop new corpora for the evaluation of spoken language understanding systems. The newly collected data are in the field of human-machine dialogue systems for tourist information in…
Semantic CompositionSpeech RecognitionSpoken Language Understanding