Assessment of L2 Oral Proficiency using Speech Large Language Models
The growing population of L2 English speakers has increased the demand for developing automatic graders for spoken language assessment (SLA). Historically, statistical models, text encoders, and self-supervised speech models have been utilised for this task. However, cascaded systems suffer from the loss of information, while E2E graders also have limitations. With the recent advancements of multi-modal large language models (LLMs), we aim to explore their potential as L2 oral proficiency graders and overcome these issues. In this work, we compare various training strategies using regression and classification targets. Our results show that speech LLMs outperform all previous competitive baselines, achieving superior performance on two datasets. Furthermore, the trained grader demonstrates strong generalisation capabilities in the cross-part or cross-task evaluation, facilitated by the audio understanding knowledge acquired during LLM pre-training.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Multi-task Pretraining for Enhancing Interpretable L2 Pronunciation Assessment
Automatic pronunciation assessment (APA) analyzes second-language (L2) learners' speech by providing fine-grained pronunciation feedback at various linguistic levels. Most existing efforts on APA typically adopt segmenta…
L2 proficiency assessment using self-supervised speech representations
There has been a growing demand for automated spoken language assessment systems in recent years. A standard pipeline for this process is to start with a speech recognition system and derive features, either hand-crafted…
speech-recognitionSpeech RecognitionSession-Level Spoken Language Assessment with a Multimodal Foundation Model via Multi-Target Learning
Spoken Language Assessment (SLA) estimates a learner's oral proficiency from spontaneous speech. The growing population of L2 English speakers has intensified the demand for reliable SLA, a critical component of Computer…
Automatic Proficiency Assessment in L2 English Learners
Second language proficiency (L2) in English is usually perceptually evaluated by English teachers or expert evaluators, with the inherent intra- and inter-rater variability. This paper explores deep learning techniques f…
Deep LearningLanguage ModelingLanguage ModellingA Novel Data Augmentation Approach for Automatic Speaking Assessment on Opinion Expressions
Automated speaking assessment (ASA) on opinion expressions is often hampered by the scarcity of labeled recordings, which restricts prompt diversity and undermines scoring reliability. To address this challenge, we propo…
Data AugmentationDiversityLanguage ModelingLanguage Modelling+6