paper-with-me

홈 › Papers

Introducing a web application for labeling, visualizing speech and correcting derived speech signals

2014-05-01 · LREC 2014 5 · Raphael Winkelmann, Georg Raess

The advent of HTML5 has sparked a great increase in interest in the web as a development platform for a variety of different research applications. Due to its ability to easily deploy software to remote clients and the recent development of standardized browser APIs, we argue that the browser has become a good platform to develop a speech labeling tool for. This paper introduces a preliminary version of an open-source client-side web application for labeling speech data, visualizing speech and segmentation information and manually correcting derived speech signals such as formant trajectories. The user interface has been designed to be as user-friendly as possible in order to make the sometimes tedious task of transcribing as easy and efficient as possible. The future integration into the next iteration of the EMU speech database management system and its general architecture will also be outlined, as the work presented here is only one of several components contributing to the future system.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Management

Similar Papers 제목 키워드 기반

SepLL: Separating Latent Class Labels from Weak Supervision Noise

2022-10-25 · Andreas Stephan, Vasiliki Kougia, Benjamin Roth

In the weakly supervised learning paradigm, labeling functions automatically assign heuristic, often noisy, labels to data samples. In this work, we provide a method for learning from weak labels by separating two types …

text-classificationText ClassificationWeakly-supervised Learning

Large Language Models based ASR Error Correction for Child Conversations

2025-05-22 · Anfeng Xu, Tiantian Feng, So Hyun Kim, Somer Bishop 외

Automatic Speech Recognition (ASR) has recently shown remarkable progress, but accurately transcribing children's speech remains a significant challenge. Recent developments in Large Language Models (LLMs) have shown pro…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Boosting Active Learning for Speech Recognition with Noisy Pseudo-labeled Samples

2020-06-19 · Jihwan Bang, Heesu Kim, Youngjoon Yoo, Jung-Woo Ha

The cost of annotating transcriptions for large speech corpora becomes a bottleneck to maximally enjoy the potential capacity of deep neural network-based automatic speech recognition models. In this paper, we present a …

Active LearningAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognition+1

Open Challenge for Correcting Errors of Speech Recognition Systems

2020-01-09 · Marek Kubis, Zygmunt Vetulani, Mikołaj Wypych, Tomasz Ziętkiewicz

The paper announces the new long-term challenge for improving the performance of automatic speech recognition systems. The goal of the challenge is to investigate methods of correcting the recognition results on the basi…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Fast Labeling and Transcription with the Speechalyzer Toolkit

2012-05-01 · LREC 2012 5 · Felix Burkhardt

We describe a software tool named “Speechalyzer” which is optimized to process large speech data sets with respect to transcription, labeling and annotation. It is implemented as a client server based framework in Java a…

Audio ClassificationBenchmarkingClassificationGeneral Classification+6