paper-with-me

홈 › Papers

End-to-End Automatic Speech Translation of Audiobooks

2018-02-12 · Alexandre Bérard, Laurent Besacier, Ali Can Kocabiyikoglu, Olivier Pietquin

We investigate end-to-end speech-to-text translation on a corpus of audiobooks specifically augmented for this task. Previous works investigated the extreme case where source language transcription is not available during learning nor decoding, but we also study a midway case where source language transcription is available at training time only. In this case, a single model is trained to decode source speech into target text in a single pass. Experimental results show that it is possible to train compact and efficient end-to-end speech translation models in this setup. We also distribute the corpus and hope that our speech translation baseline on this corpus will be challenged in the future.

📄 PDF Abstract BibTeX arXiv:1802.04200

Code (1)

alicank/Translation-Augmented-LibriSpeech-Corpus

Tasks

automatic-speech-translationSpeech-to-TextSpeech-to-Text TranslationTranslation

Similar Papers 제목 키워드 기반

LibriVoxDeEn: A Corpus for German-to-English Speech Translation and German Speech Recognition

2019-10-17 · LREC 2020 5 · Benjamin Beilharz, Xin Sun, Sariya Karimova, Stefan Riezler

We present a corpus of sentence-aligned triples of German audio, German text, and English translation, based on German audiobooks. The speech translation data consist of 110 hours of audio material aligned to over 50k pa…

Sentencespeech-recognitionSpeech RecognitionTranslation

Augmenting Librispeech with French Translations: A Multimodal Corpus for Direct Speech Translation Evaluation

2018-02-09 · LREC 2018 5 · Ali Can Kocabiyikoglu, Laurent Besacier, Olivier Kraif

Recent works in spoken language translation (SLT) have attempted to build end-to-end speech-to-text translation without using source language transcription during learning or decoding. However, while large quantities of …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine TranslationSentence+5

台語朗讀資料庫之自動切音技術應用於音文同步有聲書之建立 (Automatic Time Alignment for a Taiwanese Read Speech Corpus and its Application to Constructing Audiobooks with Text-Speech Synchronization) [In Chinese]

2012-09-01 · ROCLINGIJCLCLP 2012 9 · Wei-jay Huang, Jhih-rou Lin, Ren-Yuan Lyu, Yuang-chin Chiang 외

IESTAC: English-Italian Parallel Corpus for End-to-End Speech-to-Text Machine Translation

2020-11-01 · EMNLP (nlpbt) 2020 11 · Giuseppe Della Corte, Sara Stymne

We discuss a set of methods for the creation of IESTAC: a English-Italian speech and text parallel corpus designed for the training of end-to-end speech-to-text machine translation models and publicly released as part of…

Dynamic Time WarpingMachine TranslationSentenceSentence Embeddings+4

Large-Scale Automatic Audiobook Creation

2023-09-07 · Brendan Walsh, Mark Hamilton, Greg Newby, Xi Wang 외

An audiobook can dramatically improve a work of literature's accessibility and improve reader engagement. However, audiobooks can take hundreds of hours of human effort to create, edit, and publish. In this work, we pres…

text-to-speechText to Speech