paper-with-me

홈 › Papers

The Eloquence team submission for task 1 of MLC-SLM challenge

2025-07-25 · Lorenzo Concina, Jordi Luque, Alessio Brutti, Marco Matassoni, Yuchen Zhang arxiv

In this paper, we present our studies and experiments carried out for the task 1 of the Challenge and Workshop on Multilingual Conversational Speech Language Model (MLC-SLM), which focuses on advancing multilingual conversational speech recognition through the development of speech language models architectures. Given the increasing relevance of real-world conversational data for building robust Spoken Dialogue Systems, we explore three approaches to multilingual ASR. First, we conduct an evaluation of the official baseline to better understand its strengths and limitations, by training two projectors (linear and qformer) with different foundation models. Second we leverage the SLAM-ASR framework to train a custom multilingual linear projector. Finally we investigate the role of contrastive learning and the extended conversational context in enhancing the robustness of recognition.

📄 PDF Abstract BibTeX arXiv:2507.19308

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningSpeech Recognition

Similar Papers 제목 키워드 기반

NADI 2025: The First Multidialectal Arabic Speech Processing Shared Task

2025-09-02 · Bashar Talafha, Hawau Olamide Toyin, Peter Sullivan, AbdelRahim Elmadany 외 arxiv

We present the findings of the sixth Nuanced Arabic Dialect Identification (NADI 2025) Shared Task, which focused on Arabic speech dialect processing across three subtasks: spoken dialect identification (Subtask 1), spee…

Speech Recognition

CLUZH at SIGMORPHON 2020 Shared Task on Multilingual Grapheme-to-Phoneme Conversion

2020-07-01 · WS 2020 7 · Peter Makarov, Simon Clematide

This paper describes the submission by the team from the Institute of Computational Linguistics, Zurich University, to the Multilingual Grapheme-to-Phoneme Conversion (G2P) Task of the SIGMORPHON 2020 challenge. The subm…

Grapheme-to-Phoneme ConversionImitation Learning

NADI 2021: The Second Nuanced Arabic Dialect Identification Shared Task

2021-03-04 · EACL (WANLP) 2021 4 · Muhammad Abdul-Mageed, Chiyu Zhang, AbdelRahim Elmadany, Houda Bouamor 외

We present the findings and results of the Second Nuanced Arabic Dialect Identification Shared Task (NADI 2021). This Shared Task includes four subtasks: country-level Modern Standard Arabic (MSA) identification (Subtask…

Dialect Identification

NCUEE-NLP at MEDIQA 2021: Health Question Summarization Using PEGASUS Transformers

2021-06-01 · NAACL (BioNLP) 2021 6 · Lung-Hao Lee, Po-Han Chen, Yu-Xiang Zeng, Po-Lei Lee 외

This study describes the model design of the NCUEE-NLP system for the MEDIQA challenge at the BioNLP 2021 workshop. We use the PEGASUS transformers and fine-tune the downstream summarization task using our collected and …

I4U System Description for NIST SRE'20 CTS Challenge

2022-11-02 · Kong Aik Lee, Tomi Kinnunen, Daniele Colibro, Claudio Vair 외

This manuscript describes the I4U submission to the 2020 NIST Speaker Recognition Evaluation (SRE'20) Conversational Telephone Speech (CTS) Challenge. The I4U's submission was resulted from active collaboration among res…

Speaker Recognition