paper-with-me

홈 › Papers

Fluent Translations from Disfluent Speech in End-to-End Speech Translation

2019-06-03 · NAACL 2019 6 · Elizabeth Salesky, Matthias Sperber, Alex Waibel

Spoken language translation applications for speech suffer due to conversational speech phenomena, particularly the presence of disfluencies. With the rise of end-to-end speech translation models, processing steps such as disfluency removal that were previously an intermediate step between speech recognition and machine translation need to be incorporated into model architectures. We use a sequence-to-sequence model to translate from noisy, disfluent speech to fluent text with disfluencies removed using the recently collected `copy-edited' references for the Fisher Spanish-English dataset. We are able to directly generate fluent translations and introduce considerations about how to evaluate success on this task. This work provides a baseline for a new task, the translation of conversational speech with joint removal of disfluencies.

📄 PDF Abstract BibTeX arXiv:1906.00556

Code (0)

등록된 구현이 없습니다.

Tasks

Machine Translationspeech-recognitionSpeech RecognitionTranslation

Similar Papers 제목 키워드 기반

Generating Fluent Translations from Disfluent Text Without Access to Fluent References: IIT Bombay@IWSLT2020

2020-07-01 · WS 2020 7 · Nikhil Saini, Jyotsana Khatri, Preethi Jyothi, Pushpak Bhattacharyya

Machine translation systems perform reasonably well when the input is well-formed speech or text. Conversational speech is spontaneous and inherently consists of many disfluencies. Producing fluent translations of disflu…

DenoisingMachine TranslationTranslation

Towards Fluent Translations from Disfluent Speech

2018-11-07 · Elizabeth Salesky, Susanne Burger, Jan Niehues, Alex Waibel

When translating from speech, special consideration for conversational speech phenomena such as disfluencies is necessary. Most machine translation training data consists of well-formed written texts, causing issues when…

Machine Translationspeech-recognitionSpeech RecognitionTranslation

Fluency Over Adequacy: A Pilot Study in Measuring User Trust in Imperfect MT

2018-02-16 · WS 2018 3 · Marianna J. Martindale, Marine Carpuat

Although measuring intrinsic quality has been a key factor in the advancement of Machine Translation (MT), successfully deploying MT requires considering not just intrinsic quality but also the user experience, including…

Machine TranslationTranslation

Inclusive ASR for Disfluent Speech: Cascaded Large-Scale Self-Supervised Learning with Targeted Fine-Tuning and Data Augmentation

2024-06-14 · Dena Mujtaba, Nihar R. Mahapatra, Megan Arney, J. Scott Yaruss 외

Automatic speech recognition (ASR) systems often falter while processing stuttering-related disfluencies -- such as involuntary blocks and word repetitions -- yielding inaccurate transcripts. A critical barrier to progre…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data AugmentationSelf-Supervised Learning+2

NAIST's Machine Translation Systems for IWSLT 2020 Conversational Speech Translation Task

2020-07-01 · WS 2020 7 · Ryo Fukuda, Katsuhito Sudoh, Satoshi Nakamura

This paper describes NAIST{'}s NMT system submitted to the IWSLT 2020 conversational speech translation task. We focus on the translation disfluent speech transcripts that include ASR errors and non-grammatical utterance…

Domain AdaptationMachine TranslationNMTStyle Transfer+1