paper-with-me

Papers

NAVER LABS Europe Submission to the Instruction-following 2026 Short Track

2026-07-02 · Marcely Zanon Boito, Hemant Yadav, Jean-Luc Meunier, Ioan Calapodescu arxiv

In this paper, we describe NAVER LABS Europe's submission to the instruction-following speech processing short track at IWSLT 2026. We participate again in the constrained setting, developing systems capable of jointly performing ASR, ST, and SQA from English speech into Chinese, Italian, and German. Building on our previous submission, ranked first in last year's short track, we update our multi-stage training pipeline by replacing the speech projector with SpeechMapper, a method for learning a speech-to-LLM embedding projector using only ASR data. In addition, we introduce a synthetic SQA dataset, fakACL, composed of artificially generated scientific presentations. This dataset is built by prompting the LLM backbone, segmenting the generated talks, and synthesizing speech with SeamlessM4T-large-v2. The combination of an improved speech projection mechanism and domain-specific synthetic data allows our model to outperform last year's best short-track system, while being considerably more compact and relying on a weaker LLM backbone. This year's results place our system tied for first place in the overall short track ranking.

📄 PDF Abstract BibTeX arXiv:2607.01960

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

NAVER LABS Europe's Multilingual Speech Translation Systems for the IWSLT 2023 Low-Resource Track

2023-06-13 · Edward Gow-Smith, Alexandre Berard, Marcely Zanon Boito, Ioan Calapodescu

This paper presents NAVER LABS Europe's systems for Tamasheq-French and Quechua-Spanish speech translation in the IWSLT 2023 Low-Resource track. Our work attempts to maximize translation quality in low-resource settings …

Translation

Naver Labs Europe’s Participation in the Robustness, Chat, and Biomedical Tasks at WMT 2020

2020-11-01 · WMT (EMNLP) 2020 11 · Alexandre Berard, Ioan Calapodescu, Vassilina Nikoulina, Jerin Philip

This paper describes Naver Labs Europe’s participation in the Robustness, Chat, and Biomedical Translation tasks at WMT 2020. We propose a bidirectional German-English model that is multi-domain, robust to noise, and whi…

Language ModelingLanguage ModellingTranslation

NAVER LABS System Re-implementation for the IWSLT 2026 Instruction-Following Task

2026-07-06 · Anand Kamble, Aniket Tathe arxiv

We re-implement the NAVER LABS IWSLT 2025 instruction-following pipeline for the IWSLT 2026 Shared Task (constrained condition, short audio track), adapting it to the mandated components: SeamlessM4T-v2-large as the spee…

Naver Labs Europe (SPLADE) @ TREC Deep Learning 2022

2023-02-24 · Carlos Lassance, Stéphane Clinchant

This paper describes our participation to the 2022 TREC Deep Learning challenge. We submitted runs to all four tasks, with a focus on the full retrieval passage task. The strategy is almost the same as 2021, with first s…

Deep LearningRetrieval

Naver Labs Europe (SPLADE) @ TREC NeuCLIR 2022

2023-03-10 · Carlos Lassance, Stéphane Clinchant

This paper describes our participation in the 2022 TREC NeuCLIR challenge. We submitted runs to two out of the three languages (Farsi and Russian), with a focus on first-stage rankers and comparing mono-lingual strategie…

RetrievalTranslation