paper-with-me

홈 › Papers

A Comparative Analysis of Crowdsourced Natural Language Corpora for Spoken Dialog Systems

2016-05-01 · LREC 2016 5 · Patricia Braunger, Hansj{\"o}rg Hofmann, Steffen Werner, Maria Schmidt

Recent spoken dialog systems have been able to recognize freely spoken user input in restricted domains thanks to statistical methods in the automatic speech recognition. These methods require a high number of natural language utterances to train the speech recognition engine and to assess the quality of the system. Since human speech offers many variants associated with a single intent, a high number of user utterances have to be elicited. Developers are therefore turning to crowdsourcing to collect this data. This paper compares three different methods to elicit multiple utterances for given semantics via crowd sourcing, namely with pictures, with text and with semantic entities. Specifically, we compare the methods with regard to the number of valid data and linguistic variance, whereby a quantitative and qualitative approach is proposed. In our study, the method with text led to a high variance in the utterances and a relatively low rate of invalid data.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognitionvalid

Similar Papers 제목 키워드 기반

Efficiency and Effectiveness of LLM-Based Summarization of Evidence in Crowdsourced Fact-Checking

2025-01-30 · Kevin Roitero, Dustin Wright, Michael Soprano, Isabelle Augenstein 외

Evaluating the truthfulness of online content is critical for combating misinformation. This study examines the efficiency and effectiveness of crowdsourced truthfulness assessments through a comparative analysis of two …

Fact CheckingLanguage ModelingLanguage ModellingLarge Language Model+1

Letting the Data Speak: Extracting Keywords from Crowdsourced Collections with AI

2026-07-10 · Miguel Arana-Catania, Catherine Conisbee, Matthew Kidd arxiv

Identifying and assigning keywords at scale is a technical, practical, and ethical challenge for crowdsourced collections. This article reports the findings of the "Extracting Keywords from Crowdsourced Collections" proj…

Keyword Extraction

Crowdsourced Multilingual Speech Intelligibility Testing

2024-03-21 · Laura Lechler, Kamil Wojcicki

With the advent of generative audio features, there is an increasing need for rapid evaluation of their impact on speech intelligibility. Beyond the existing laboratory measures, which are expensive and do not scale well…

Speech Intelligibility Evaluation

Comparative Opinion Mining: A Review

2017-12-24 · Kasturi Dewi Varathan, Anastasia Giachanou, Fabio Crestani

Opinion mining refers to the use of natural language processing, text analysis and computational linguistics to identify and extract subjective information in textual material. Opinion mining, also known as sentiment ana…

Opinion MiningSentiment Analysis

Justifying Social-Choice Mechanism Outcome for Improving Participant Satisfaction

2022-05-24 · Sharadhi Alape Suryanarayana, David Sarne, Sarit Kraus

In many social-choice mechanisms the resulting choice is not the most preferred one for some of the participants, thus the need for methods to justify the choice made in a way that improves the acceptance and satisfactio…