paper-with-me

홈 › Papers

Advanced Rich Transcription System for Estonian Speech

2019-01-11 · Tanel Alumäe, Ottokar Tilk, Asadullah

This paper describes the current TT\"U speech transcription system for Estonian speech. The system is designed to handle semi-spontaneous speech, such as broadcast conversations, lecture recordings and interviews recorded in diverse acoustic conditions. The system is based on the Kaldi toolkit. Multi-condition training using background noise profiles extracted automatically from untranscribed data is used to improve the robustness of the system. Out-of-vocabulary words are recovered using a phoneme n-gram based decoding subgraph and a FST-based phoneme-to-grapheme model. The system achieves a word error rate of 8.1% on a test set of broadcast conversations. The system also performs punctuation recovery and speaker identification. Speaker identification models are trained using a recently proposed weakly supervised training method.

📄 PDF Abstract BibTeX arXiv:1901.03601

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Identification

Similar Papers 제목 키워드 기반

End-to-End Rich Transcription-Style Automatic Speech Recognition with Semi-Supervised Learning

2021-07-07 · Tomohiro Tanaka, Ryo Masumura, Mana Ihori, Akihiko Takashima 외

We propose a semi-supervised learning method for building end-to-end rich transcription-style automatic speech recognition (RT-ASR) systems from small-scale rich transcription-style and large-scale common transcription-s…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Finetuning End-to-End Models for Estonian Conversational Spoken Language Translation

2024-07-04 · Tiia Sildam, Andra Velve, Tanel Alumäe

This paper investigates the finetuning of end-to-end models for bidirectional Estonian-English and Estonian-Russian conversational speech-to-text translation. Due to the limited availability of speech translation data fo…

Machine Translationspeech-recognitionSpeech RecognitionSpeech-to-Text+2

Orthographic Transcription: which enrichment is required for phonetization?

2012-05-01 · LREC 2012 5 · Brigitte Bigi, Pauline Péri, Roxane Bertrand

This paper addresses the problem of the enrichment of transcriptions in the perspective of an automatic phonetization. Phonetization is the process of representing sounds with phonetic signs. There are two general ways t…

Neural Speech Synthesis for Estonian

2020-10-06 · Liisa Rätsep, Liisi Piits, Hille Pajupuu, Indrek Hein 외

This technical report describes the results of a collaboration between the NLP research group at the University of Tartu and the Institute of Estonian Language on improving neural speech synthesis for Estonian. The repor…

SentenceSpeech Synthesistext-to-speechText to Speech

MTee: Open Machine Translation Platform for Estonian Government

2022-06-01 · EAMT 2022 6 · Toms Bergmanis, Marcis Pinnis, Roberts Rozis, Jānis Šlapiņš 외

We present the MTee project - a research initiative funded via an Estonian public procurement to develop machine translation technology that is open-source and free of charge. The MTee project delivered an open-source pl…

Document TranslationGrammatical Error CorrectionMachine TranslationTranslation