paper-with-me

Papers

NAVER LABS System Re-implementation for the IWSLT 2026 Instruction-Following Task

2026-07-06 · Anand Kamble, Aniket Tathe arxiv

We re-implement the NAVER LABS IWSLT 2025 instruction-following pipeline for the IWSLT 2026 Shared Task (constrained condition, short audio track), adapting it to the mandated components: SeamlessM4T-v2-large as the speech encoder and Qwen3-4B-Instruct as the LLM backbone. The three-stage approach projector alignment, text-only LoRA pre-training, and multimodal merging is preserved from the original design. We additionally construct 100k synthetic instruction-following examples across ten speech-centric task types (10k per task) from the provided corpora, suitable for further Stage 3 fine-tuning. Our primary model achieves COMET 0.781 on EN-ZH speech translation and BERTScore-F1 0.346 on English SQA on the MCIF benchmark.

📄 PDF Abstract BibTeX arXiv:2607.05623

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

NAVER LABS Europe's Multilingual Speech Translation Systems for the IWSLT 2023 Low-Resource Track

2023-06-13 · Edward Gow-Smith, Alexandre Berard, Marcely Zanon Boito, Ioan Calapodescu

This paper presents NAVER LABS Europe's systems for Tamasheq-French and Quechua-Spanish speech translation in the IWSLT 2023 Low-Resource track. Our work attempts to maximize translation quality in low-resource settings …

Translation

NAVER LABS Europe Submission to the Instruction-following 2026 Short Track

2026-07-02 · Marcely Zanon Boito, Hemant Yadav, Jean-Luc Meunier, Ioan Calapodescu arxiv

In this paper, we describe NAVER LABS Europe's submission to the instruction-following speech processing short track at IWSLT 2026. We participate again in the constrained setting, developing systems capable of jointly p…

DiDi Labs' End-to-end System for the IWSLT 2020 Offline Speech TranslationTask

2020-07-01 · WS 2020 7 · Arkady Arkhangorodsky, Yiqi Huang, Amittai Axelrod

This paper describes the system that was submitted by DiDi Labs to the offline speech translation task for IWSLT 2020. We trained an end-to-end system that translates audio from English TED talks to German text, without …

Translation

Naver Labs Europe’s Participation in the Robustness, Chat, and Biomedical Tasks at WMT 2020

2020-11-01 · WMT (EMNLP) 2020 11 · Alexandre Berard, Ioan Calapodescu, Vassilina Nikoulina, Jerin Philip

This paper describes Naver Labs Europe’s participation in the Robustness, Chat, and Biomedical Translation tasks at WMT 2020. We propose a bidirectional German-English model that is multi-domain, robust to noise, and whi…

Language ModelingLanguage ModellingTranslation

Octanove Labs' Japanese-Chinese Open Domain Translation System

2020-07-01 · WS 2020 7 · Masato Hagiwara

This paper describes Octanove Labs{'} submission to the IWSLT 2020 open domain translation challenge. In order to build a high-quality Japanese-Chinese neural machine translation (NMT) system, we use a combination of 1) …

Machine TranslationNMTTranslation