Segmentation Strategies for Streaming Speech Translation
Code (0)
등록된 구현이 없습니다.
Tasks
ChunkingLanguage ModellingMachine TranslationSegmentationSpeech RecognitionTranslationSimilar Papers 제목 키워드 기반
End-to-End Simultaneous Speech Translation with Differentiable Segmentation
End-to-end simultaneous speech translation (SimulST) outputs translation while receiving the streaming speech inputs (a.k.a. streaming speech translation), and hence needs to segment the speech inputs and then translate …
SegmentationTranslationDirect Segmentation Models for Streaming Speech Translation
The cascade approach to Speech Translation (ST) is based on a pipeline that concatenates an Automatic Speech Recognition (ASR) system followed by a Machine Translation (MT) system. These systems are usually connected by …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine TranslationSegmentation+3Learning Adaptive Segmentation Policy for End-to-End Simultaneous Translation
End-to-end simultaneous speech-to-text translation aims to directly perform translation from streaming source speech to target text with high translation quality and low latency. A typical simultaneous translation (ST) s…
SegmentationSimultaneous Speech-to-Text TranslationSpeech-to-TextSpeech-to-Text Translation+1Learning When to Translate for Streaming Speech
How to find proper moments to generate partial sentence translation given a streaming speech input? Existing approaches waiting-and-translating for a fixed duration often break the acoustic units in speech, since the bou…
DecoderSentenceSpeech-to-Text TranslationTranslationStreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model
Streaming speech translation (StreamST) requires determining appropriate timing, known as policy, to generate translations while continuously receiving source speech inputs, balancing low latency with high translation qu…