paper-with-me

홈 › Papers

JPS-daprinfo: A Dataset for Japanese Dialog Act Analysis and People-related Information Detection

2021-03-06 · Changzeng Fu

We conducted a labeling work on a spoken Japanese dataset (I-JAS) for the text classification, which contains 50 interview dialogues of two-way Japanese conversation that discuss the participants' past present and future. Each dialogue is 30 minutes long. From this dataset, we selected the interview dialogues of native Japanese speakers as the samples. Given the dataset, we annotated sentences with 13 labels. The labeling work was conducted by native Japanese speakers who have experiences with data annotation. The total amount of the annotated samples is 20130.

📄 PDF Abstract BibTeX arXiv:2103.11786

Code (1)

CZFuChason/JPS-daprinfo 공식 구현

Tasks

text-classificationText Classification

Similar Papers 제목 키워드 기반

Speech Corpus Spoken by Young-old, Old-old and Oldest-old Japanese

2016-05-01 · LREC 2016 5 · Yurie Iribe, Norihide Kitaoka, Shuhei Segawa

We have constructed a new speech data corpus, using the utterances of 100 elderly Japanese people, to improve speech recognition accuracy of the speech of older people. Humanoid robots are being developed for use in elde…

speech-recognitionSpeech Recognition

Empirical Analysis of Training Strategies of Transformer-based Japanese Chit-chat Systems

2021-09-11 · Hiroaki Sugiyama, Masahiro Mizukami, Tsunehiro Arimoto, Hiromi Narimatsu 외

In recent years, several high-performance conversational systems have been proposed based on the Transformer encoder-decoder model. Although previous studies analyzed the effects of the model parameters and the decoding …

Decoder

Empirical Analysis of Training Strategies of Transformer-based Japanese Chit-chat Systems

2021-11-16 · ACL ARR November 2021 11 · Anonymous

In recent years, several high-performance conversational systems have been proposed based on the Transformer encoder-decoder model. Although previous studies analyzed the effects of the model parameters and the decoding …

Decoder

JMultiWOZ: A Large-Scale Japanese Multi-Domain Task-Oriented Dialogue Dataset

2024-03-26 · Atsumoto Ohashi, Ryu Hirai, Shinya Iizuka, Ryuichiro Higashinaka

Dialogue datasets are crucial for deep learning-based task-oriented dialogue system research. While numerous English language multi-domain task-oriented dialogue datasets have been developed and contributed to significan…

Dialogue State TrackingLanguage ModelingLanguage ModellingLarge Language Model+2

EmplifAI: a Fine-grained Dataset for Japanese Empathetic Medical Dialogues in 28 Emotion Labels

2026-01-15 · Wan Jou She, Lis Kanashiro Pereira, Fei Cheng, Sakiko Yahata 외 arxiv

This paper introduces EmplifAI, a Japanese empathetic dialogue dataset designed to support patients coping with chronic medical conditions. They often experience a wide range of positive and negative emotions (e.g., hope…