paper-with-me

홈 › Papers

Robust Translation of French Live Speech Transcripts

2022-09-01 · AMTA 2022 9 · Elise Bertin-Lemée, Guillaume Klein, Josep Crego, Jean Senellart

Despite a narrowed performance gap with direct approaches, cascade solutions, involving automatic speech recognition (ASR) and machine translation (MT) are still largely employed in speech translation (ST). Direct approaches employing a single model to translate the input speech signal suffer from the critical bottleneck of data scarcity. In addition, multiple industry applications display speech transcripts alongside translations, making cascade approaches more realistic and practical. In the context of cascaded simultaneous ST, we propose several solutions to adapt a neural MT network to take as input the transcripts output by an ASR system. Adaptation is achieved by enriching speech transcripts and MT data sets so that they more closely resemble each other, thereby improving the system robustness to error propagation and enhancing result legibility for humans. We address aspects such as sentence boundaries, capitalisation, punctuation, hesitations, repetitions, homophones, etc. while taking into account the low latency requirement of simultaneous ST systems.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine TranslationSentencespeech-recognitionSpeech RecognitionTranslation

Similar Papers 제목 키워드 기반

Microsoft Speech Language Translation (MSLT) Corpus: The IWSLT 2016 release for English, French and German

2016-12-01 · IWSLT 2016 12 · Christian Federmann, William D. Lewis

We describe the Microsoft Speech Language Translation (MSLT) corpus, which was created in order to evaluate end-to-end conversational speech translation quality. The corpus was created from actual conversations over Skyp…

Machine Translationspeech-recognitionSpeech RecognitionTranslation

The UMD Machine Translation Systems at IWSLT 2016: English-to-French Translation of Speech Transcripts

2016-12-01 · IWSLT 2016 12 · Xing Niu, Marine Carpuat

We describe the University of Maryland machine translation system submitted to the IWSLT 2016 Microsoft Speech Language Translation (MSLT) English-French task. Our main finding is that translating conversation transcript…

Machine TranslationTranslation

Listen and Translate: A Proof of Concept for End-to-End Speech-to-Text Translation

2016-12-06 · Alexandre Berard, Olivier Pietquin, Christophe Servan, Laurent Besacier

This paper proposes a first attempt to build an end-to-end speech-to-text translation system, which does not use source language transcription during learning or decoding. We propose a model for direct speech-to-text tra…

Speech-to-TextSpeech-to-Text TranslationTranslation

SkinAugment: Auto-Encoding Speaker Conversions for Automatic Speech Translation

2020-02-27 · Arya D. McCarthy, Liezl Puzon, Juan Pino

We propose autoencoding speaker conversion for training data augmentation in automatic speech translation. This technique directly transforms an audio sequence, resulting in audio synthesized to resemble another speaker'…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)automatic-speech-translationData Augmentation+4

HK-LegiCoST: Leveraging Non-Verbatim Transcripts for Speech Translation

2023-06-20 · Cihan Xiao, Henry Li Xinyuan, Jinyi Yang, Dongji Gao 외

We introduce HK-LegiCoST, a new three-way parallel corpus of Cantonese-English translations, containing 600+ hours of Cantonese audio, its standard traditional Chinese transcript, and English translation, segmented and a…

Cross-corpusSentencespeech-recognitionSpeech Recognition+1