paper-with-me

홈 › Papers

A Case Study on Filtering for End-to-End Speech Translation

2024-02-02 · Md Mahfuz ibn Alam, Antonios Anastasopoulos

It is relatively easy to mine a large parallel corpus for any machine learning task, such as speech-to-text or speech-to-speech translation. Although these mined corpora are large in volume, their quality is questionable. This work shows that the simplest filtering technique can trim down these big, noisy datasets to a more manageable, clean dataset. We also show that using this clean dataset can improve the model's performance, as in the case of the multilingual-to-English Speech Translation (ST) model, where, on average, we obtain a 4.65 BLEU score improvement.

📄 PDF Abstract BibTeX arXiv:2402.01945

Code (0)

등록된 구현이 없습니다.

Tasks

Speech-to-Speech TranslationSpeech-to-TextTranslation

Similar Papers 제목 키워드 기반

VAKTA-SETU: A Speech-to-Speech Machine Translation Service in Select Indic Languages

2023-05-21 · Shivam Mhaskar, Vineet Bhat, Akshay Batheja, Sourabh Deoghare 외

In this work, we present our deployment-ready Speech-to-Speech Machine Translation (SSMT) system for English-Hindi, English-Marathi, and Hindi-Marathi language pairs. We develop the SSMT system by cascading Automatic Spe…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine Translationspeech-recognition+6

End-to-End Automatic Speech Translation of Audiobooks

2018-02-12 · Alexandre Bérard, Laurent Besacier, Ali Can Kocabiyikoglu, Olivier Pietquin

We investigate end-to-end speech-to-text translation on a corpus of audiobooks specifically augmented for this task. Previous works investigated the extreme case where source language transcription is not available durin…

automatic-speech-translationSpeech-to-TextSpeech-to-Text TranslationTranslation

Leveraging Audio-LLMs to Filter Speech-to-Speech Training Data

2026-06-11 · Qixu Chen, Satoshi Nakamura arxiv

Large-scale mined corpora provide abundant training data for end-to-end speech-to-speech translation (S2ST) but may contain noise, misalignment, and semantic errors. Filtering noisy data is crucial to maintain robust spe…

Speech-to-Speech Translation

A case study on using speech-to-translation alignments for language documentation

2017-02-14 · WS 2017 3 · Antonios Anastasopoulos, David Chiang

For many low-resource or endangered languages, spoken language resources are more likely to be annotated with translations than with transcriptions. Recent work exploits such annotations to produce speech-to-translation …

speech-recognitionSpeech RecognitionTranslation

Speech-to-Speech Translation For A Real-world Unwritten Language

2022-11-11 · arXiv 2022 10 · Peng-Jen Chen, Kevin Tran, Yilin Yang, Jingfei Du 외

We study speech-to-speech translation (S2ST) that translates speech from one language into another language and focuses on building systems to support languages without standard text writing systems. We use English-Taiwa…

Speech-to-Speech TranslationTranslation