paper-with-me

홈 › Papers

Developing automatic verbatim transcripts for international multilingual meetings: an end-to-end solution

2023-09-27 · Akshat Dewan, Michal Ziemski, Henri Meylan, Lorenzo Concina, Bruno Pouliquen

This paper presents an end-to-end solution for the creation of fully automated conference meeting transcripts and their machine translations into various languages. This tool has been developed at the World Intellectual Property Organization (WIPO) using in-house developed speech-to-text (S2T) and machine translation (MT) components. Beyond describing data collection and fine-tuning, resulting in a highly customized and robust system, this paper describes the architecture and evolution of the technical components as well as highlights the business impact and benefits from the user side. We also point out particular challenges in the evolution and adoption of the system and how the new approach created a new product and replaced existing established workflows in conference management documentation.

📄 PDF Abstract BibTeX arXiv:2309.15609

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationManagementSpeech-to-Text

Similar Papers 제목 키워드 기반

Europarl-ASR: A Large Corpus of Parliamentary Debates for Streaming ASR Benchmarking and Speech Data Filtering/Verbatimization

2021-08-30 · Interspeech 2021 8 · Gonçal V. Garcés Díaz-Munío, Joan-Albert Silvestre-Cerdà, Javier Jorge, Adrià Giménez Pastor 외

We introduce Europarl-ASR, a large speech and text corpus of parliamentary debates including 1 300 hours of transcribed speeches and 70 million tokens of text in English extracted from European Parliament sessions. The t…

BenchmarkingData AugmentationSpeech Recognition

Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling

2025-02-05 · Jakob Poncelet, Hugo Van hamme

The recent advancement of speech recognition technology has been driven by large-scale datasets and attention-based architectures, but many challenges still remain, especially for low-resource languages and dialects. Thi…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

MONAH: Multi-Modal Narratives for Humans to analyze conversations

2021-01-18 · EACL 2021 2 · Joshua Y. Kim, Greyson Y. Kim, Chunfeng Liu, Rafael A. Calvo 외

In conversational analyses, humans manually weave multimodal information into the transcripts, which is significantly time-consuming. We introduce a system that automatically expands the verbatim transcripts of video-rec…

Emotion Recognition in ConversationFeature EngineeringText Generation

Automatic Detection of Everyday Social Behaviours and Environments from Verbatim Transcripts of Daily Conversations

2019-07-22 · Kristina Y. Yordanova, Burcu Demiray, Matthias R. Mehl, Mike Martin

Coding in social sciences is a process that involves the categorisation of qualitative or quantitative data in order to facilitate further analysis. Coding is usually a manual process that involves a lot of effort and ti…

Quantification of stylistic differences in human- and ASR-produced transcripts of African American English

2024-09-04 · Annika Heuser, Tyler Kendall, Miguel Del Rio, Quinten McNamara 외

Common measures of accuracy used to assess the performance of automatic speech recognition (ASR) systems, as well as human transcribers, conflate multiple sources of error. Stylistic differences, such as verbatim vs non-…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition