paper-with-me

Papers

SimulSeamless: FBK at IWSLT 2024 Simultaneous Speech Translation

2024-06-20 · Sara Papi, Marco Gaido, Matteo Negri, Luisa Bentivogli

This paper describes the FBK's participation in the Simultaneous Translation Evaluation Campaign at IWSLT 2024. For this year's submission in the speech-to-text translation (ST) sub-track, we propose SimulSeamless, which is realized by combining AlignAtt and SeamlessM4T in its medium configuration. The SeamlessM4T model is used "off-the-shelf" and its simultaneous inference is enabled through the adoption of AlignAtt, a SimulST policy based on cross-attention that can be applied without any retraining or adaptation of the underlying model for the simultaneous task. We participated in all the Shared Task languages (English->{German, Japanese, Chinese}, and Czech->English), achieving acceptable or even better results compared to last year's submissions. SimulSeamless, covering more than 143 source languages and 200 target languages, is released at: https://github.com/hlt-mt/FBK-fairseq/.

📄 PDF Abstract BibTeX arXiv:2406.14177

Code (1)

hlt-mt/fbk-fairseq 공식 구현 pytorch

Tasks

Speech-to-TextSpeech-to-Text TranslationTranslation

Similar Papers 제목 키워드 기반

MLLP-VRAIN UPV systems for the IWSLT 2022 Simultaneous Speech Translation and Speech-to-Speech Translation tasks

2022-05-01 · IWSLT (ACL) 2022 5 · Javier Iranzo-Sánchez, Javier Jorge Cano, Alejandro Pérez-González-de-Martos, Adrián Giménez Pastor 외

This work describes the participation of the MLLP-VRAIN research group in the two shared tasks of the IWSLT 2022 conference: Simultaneous Speech Translation and Speech-to-Speech Translation. We present our streaming-read…

Simultaneous Speech-to-Text TranslationSpeech-to-Speech TranslationTranslation

NAIST Simultaneous Speech-to-Text Translation System for IWSLT 2022

2022-05-01 · IWSLT (ACL) 2022 5 · Ryo Fukuda, Yuka Ko, Yasumasa Kano, Kosuke Doi 외

This paper describes NAIST’s simultaneous speech translation systems developed for IWSLT 2022 Evaluation Campaign. We participated the speech-to-speech track for English-to-German and English-to-Japanese. Our primary sub…

SegmentationSimultaneous Speech-to-Text TranslationSpeech-to-TextSpeech-to-Text Translation+1

The Volctrans Neural Speech Translation System for IWSLT 2021

2021-05-16 · ACL (IWSLT) 2021 8 · Chengqi Zhao, Zhicheng Liu, Jian Tong, Tao Wang 외

This paper describes the systems submitted to IWSLT 2021 by the Volctrans team. We participate in the offline speech translation and text-to-text simultaneous translation tracks. For offline speech translation, our best …

Translation

A Pocket Offline Model for Simultaneous Speech Translation as CUNI Submission to IWSLT 2026

2026-06-02 · Aziz Sharipov Ortega, Dominik Macháček arxiv

We implement simultaneous translation capability with the offline direct speech-to-text translation model Canary, using the state-of-the-art policy AlignAtt, and submit it to IWSLT 2026 Simultaneous Speech Translation Sh…

Speech-to-Text Translation

FINDINGS OF THE IWSLT 2021 EVALUATION CAMPAIGN

2021-08-01 · ACL (IWSLT) 2021 8 · Antonios Anastasopoulos, Ondřej Bojar, Jacob Bremerman, Roldano Cattoni 외

The evaluation campaign of the International Conference on Spoken Language Translation (IWSLT 2021) featured this year four shared tasks: (i) Simultaneous speech translation, (ii) Offline speech translation, (iii) Multil…

Translation