paper-with-me

홈 › Papers

SMASH Corpus: A Spontaneous Speech Corpus Recording Third-person Audio Commentaries on Gameplay

2020-05-01 · LREC 2020 5 · Yuki Saito, Shinnosuke Takamichi, Hiroshi Saruwatari

Developing a spontaneous speech corpus would be beneficial for spoken language processing and understanding. We present a speech corpus named the SMASH corpus, which includes spontaneous speech of two Japanese male commentators that made third-person audio commentaries during the gameplay of a fighting game. Each commentator ad-libbed while watching the gameplay with various topics covering not only explanations of each moment to convey the information on the fight but also comments to entertain listeners. We made transcriptions and topic tags as annotations on the recorded commentaries with our two-step method. We first made automatic and manual transcriptions of the commentaries and then manually annotated the topic tags. This paper describes how we constructed the SMASH corpus and reports some results of the annotations.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CoRuSS - a New Prosodically Annotated Corpus of Russian Spontaneous Speech

2016-05-01 · LREC 2016 5 · Tatiana Kachkovskaia, Daniil Kocharov, Pavel Skrelin, Nina Volskaya

This paper describes speech data recording, processing and annotation of a new speech corpus CoRuSS (Corpus of Russian Spontaneous Speech), which is based on connected communicative speech recorded from 60 native Russian…

Construction of a Large-scale Japanese ASR Corpus on TV Recordings

2021-03-26 · Shintaro Ando, Hiromasa Fujihara

This paper presents a new large-scale Japanese speech corpus for training automatic speech recognition (ASR) systems. This corpus contains over 2,000 hours of speech with transcripts built on Japanese TV recordings and t…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

A Corpus of Read and Spontaneous Upper Saxon German Speech for ASR Evaluation

2016-05-01 · LREC 2016 5 · Robert Herms, Laura Seelig, Stefanie M{\"u}nch, Maximilian Eibl

In this Paper we present a corpus named SXUCorpus which contains read and spontaneous speech of the Upper Saxon German dialect. The data has been collected from eight archives of local television stations located in the …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+2

GRASS: the Graz corpus of Read And Spontaneous Speech

2014-05-01 · LREC 2014 5 · Barbara Schuppler, Martin Hagmueller, Juan A. Morales-Cordovilla, Hannes Pessentheiner

This paper provides a description of the preparation, the speakers, the recordings, and the creation of the orthographic transcriptions of the first large scale speech database for Austrian German. It contains approximat…

Speech Recognition

SaSLaW: Dialogue Speech Corpus with Audio-visual Egocentric Information Toward Environment-adaptive Dialogue Speech Synthesis

2024-08-13 · Osamu Take, Shinnosuke Takamichi, Kentaro Seki, Yoshiaki Bando 외

This paper presents SaSLaW, a spontaneous dialogue speech corpus containing synchronous recordings of what speakers speak, listen to, and watch. Humans consider the diverse environmental factors and then control the feat…

Speech SynthesisSpoken Dialogue Systemstext-to-speechText to Speech