paper-with-me

홈 › Papers

JSSS: free Japanese speech corpus for summarization and simplification

2020-10-05

In this paper, we construct a new Japanese speech corpus for speech-based summarization and simplification, "JSSS" (pronounced "j-triple-s"). Given the success of reading-style speech synthesis from short-form sentences, we aim to design more difficult tasks for delivering information to humans. Our corpus contains voices recorded for two tasks that have a role in providing information under constraints: duration-constrained text-to-speech summarization and speaking-style simplification. It also contains utterances of long-form sentences as an optional task. This paper describes how we designed the corpus, which is available on our project page.

📄 PDF Abstract BibTeX arXiv:2010.01793

Code (1)

tarepan/jsss pytorch

Tasks

FormSpeech Synthesistext-to-speechText to Speech

Similar Papers 제목 키워드 기반

JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis

2017-10-28 · Ryosuke Sonobe, Shinnosuke Takamichi, Hiroshi Saruwatari

Thanks to improvements in machine learning techniques including deep learning, a free large-scale speech corpus that can be shared between academic institutions and commercial companies has an important role. However, su…

BIG-bench Machine LearningSpeech Synthesis

The Kyutech corpus and topic segmentation using a combined method

2016-12-01 · WS 2016 12 · Takashi Yamamura, Kazutaka Shimada, Shintaro Kawahara

Summarization of multi-party conversation is one of the important tasks in natural language processing. In this paper, we explain a Japanese corpus and a topic segmentation task. To the best of our knowledge, the corpus …

Decision MakingSegmentation

Construction of a Large-scale Japanese ASR Corpus on TV Recordings

2021-03-26 · Shintaro Ando, Hiromasa Fujihara

This paper presents a new large-scale Japanese speech corpus for training automatic speech recognition (ASR) systems. This corpus contains over 2,000 hours of speech with transcripts built on Japanese TV recordings and t…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Speech Corpus Spoken by Young-old, Old-old and Oldest-old Japanese

2016-05-01 · LREC 2016 5 · Yurie Iribe, Norihide Kitaoka, Shuhei Segawa

We have constructed a new speech data corpus, using the utterances of 100 elderly Japanese people, to improve speech recognition accuracy of the speech of older people. Humanoid robots are being developed for use in elde…

speech-recognitionSpeech Recognition

CPJD Corpus: Crowdsourced Parallel Speech Corpus of Japanese Dialects

2018-05-01 · LREC 2018 5 · Shinnosuke Takamichi, Hiroshi Saruwatari
Machine TranslationSpeech RecognitionSpeech Synthesis