paper-with-me

홈 › Papers

Enabling robust and fluid spoken dialogue with cognitively impaired users

2017-08-01 · WS 2017 8 · Ramin Yaghoubzadeh, Stefan Kopp

We present the flexdiam dialogue management architecture, which was developed in a series of projects dedicated to tailoring spoken interaction to the needs of users with cognitive impairments in an everyday assistive domain, using a multimodal front-end. This hybrid DM architecture affords incremental processing of uncertain input, a flexible, mixed-initiative information grounding process that can be adapted to users{'} cognitive capacities and interactive idiosyncrasies, and generic mechanisms that foster transitions in the joint discourse state that are understandable and controllable by those users, in order to effect a robust interaction for users with varying capacities.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Dialogue ManagementManagementSpeech Recognition

Similar Papers 제목 키워드 기반

Towards human-like spoken dialogue generation between AI agents from written dialogue

2023-10-02 · Kentaro Mitsui, Yukiya Hono, Kei Sawada

The advent of large language models (LLMs) has made it possible to generate natural written dialogues between two agents. However, generating human-like spoken dialogues from these written dialogues remains challenging. …

Dialogue Generationtext-to-speechText to Speech

Real-Time Textless Dialogue Generation

2025-01-08 · Long Mai, Julie Carson-Berndsen

Recent advancements in large language models (LLMs) have led to significant progress in text-based dialogue systems. These systems can now generate high-quality responses that are accurate and coherent across a wide rang…

Dialogue GenerationRhythmSpoken Dialogue Systems

Generative Spoken Dialogue Language Modeling

2022-03-30 · Tu Anh Nguyen, Eugene Kharitonov, Jade Copet, Yossi Adi 외

We introduce dGSLM, the first "textless" model able to generate audio samples of naturalistic spoken dialogues. It uses recent work on unsupervised spoken unit discovery coupled with a dual-tower transformer architecture…

Language ModelingLanguage Modelling

NTPP: Generative Speech Language Modeling for Dual-Channel Spoken Dialogue via Next-Token-Pair Prediction

2025-06-01 · Qichao Wang, Ziqiao Meng, Wenqian Cui, Yifei Zhang 외

Inspired by the impressive capabilities of GPT-4o, there is growing interest in enabling speech language models (SLMs) to engage in natural, fluid spoken interactions with humans. Recent advancements have led to the deve…

DecoderLanguage ModelingLanguage Modelling

WavChat: A Survey of Spoken Dialogue Models

2024-11-15 · Shengpeng Ji, Yifu Chen, Minghui Fang, Jialong Zuo 외

Recent advancements in spoken dialogue models, exemplified by systems like GPT-4o, have captured significant attention in the speech domain. Compared to traditional three-tier cascaded spoken dialogue models that compris…

speech-recognitionSpeech RecognitionSpoken Dialogue SystemsSurvey+2