paper-with-me

홈 › Papers

EmoNews: A Spoken Dialogue System for Expressive News Conversations

2025-06-16 · Ryuki Matsuura, Shikhar Bharadwaj, Jiarui Liu, Dhatchi Kunde Govindarajan

We develop a task-oriented spoken dialogue system (SDS) that regulates emotional speech based on contextual cues to enable more empathetic news conversations. Despite advancements in emotional text-to-speech (TTS) techniques, task-oriented emotional SDSs remain underexplored due to the compartmentalized nature of SDS and emotional TTS research, as well as the lack of standardized evaluation metrics for social goals. We address these challenges by developing an emotional SDS for news conversations that utilizes a large language model (LLM)-based sentiment analyzer to identify appropriate emotions and PromptTTS to synthesize context-appropriate emotional speech. We also propose subjective evaluation scale for emotional SDSs and judge the emotion regulation performance of the proposed and baseline systems. Experiments showed that our emotional SDS outperformed a baseline system in terms of the emotion regulation and engagement. These results suggest the critical role of speech emotion for more engaging conversations. All our source code is open-sourced at https://github.com/dhatchi711/espnet-emotional-news/tree/emo-sds/egs2/emo_news_sds/sds1

📄 PDF Abstract BibTeX arXiv:2506.13894

Code (1)

dhatchi711/espnet-emotional-news 공식 구현 pytorch

Tasks

Language ModelingLanguage ModellingLarge Language Modeltext-to-speechText to Speech

Similar Papers 제목 키워드 기반

WavAlign: Enhancing Intelligence and Expressiveness in Spoken Dialogue Models via Adaptive Hybrid Post-Training

2026-04-16 · Yifu Chen, Shengpeng Ji, Qian Chen, Tianle Liang 외 arxiv

End-to-end spoken dialogue models have garnered significant attention because they offer a higher potential ceiling in expressiveness and perceptual ability than cascaded systems. However, the intelligence and expressive…

Reinforcement Learning

Linear Semantic Segmentation for Low-Resource Spoken Dialects

2026-05-07 · Kirill Chirkunov, Younes Samih, Abed Alhakim Freihat, Hanan Aldarmaki arxiv

Semantic segmentation is a core component of discourse analysis, yet existing models are primarily developed and evaluated on high-resource written text, limiting their effectiveness on low-resource spoken varieties. In …

Semantic Segmentation

LD-SDS: Towards an Expressive Spoken Dialogue System based on Linked-Data

2017-10-09 · Alexandros Papangelis, Panagiotis Papadakos, Margarita Kotti, Yannis Stylianou 외

In this work we discuss the related challenges and describe an approach towards the fusion of state-of-the-art technologies from the Spoken Dialogue Systems (SDS) and the Semantic Web and Information Retrieval domains. W…

Conversational SearchInformation RetrievalRetrievalslot-filling+2

RoleBreak: Benchmarking Long-Horizon Role-Playing Robustness in Spoken Dialogue

2026-09-15 · Yuqi Wang, Fengyuan Liu, Haochen Luo, Zhiqi Yu 외 arxiv

Speech-to-speech dialogue models increasingly support persona control, yet existing spoken role-playing benchmarks remain largely character-centric and short-horizon. This leaves open whether spoken dialogue models can s…

Neural Dialogue State Tracking with Temporally Expressive Networks

2020-09-16 · Findings of the Association for Computational Linguistics 2020 · Junfan Chen, Richong Zhang, Yongyi Mao, Jie Xu

Dialogue state tracking (DST) is an important part of a spoken dialogue system. Existing DST models either ignore temporal feature dependencies across dialogue turns or fail to explicitly model temporal state dependencie…

Dialogue State Tracking