paper-with-me

홈 › Papers

Exploring Textual and Speech information in Dialogue Act Classification with Speaker Domain Adaptation

2018-10-17 · ALTA 2018 12 · Xuanli He, Quan Hung Tran, William Havard, Laurent Besacier, Ingrid Zukerman, Gholamreza Haffari

In spite of the recent success of Dialogue Act (DA) classification, the majority of prior works focus on text-based classification with oracle transcriptions, i.e. human transcriptions, instead of Automatic Speech Recognition (ASR)'s transcriptions. In spoken dialog systems, however, the agent would only have access to noisy ASR transcriptions, which may further suffer performance degradation due to domain shift. In this paper, we explore the effectiveness of using both acoustic and textual signals, either oracle or ASR transcriptions, and investigate speaker domain adaptation for DA classification. Our multimodal model proves to be superior to the unimodal models, particularly when the oracle transcriptions are not available. We also propose an effective method for speaker domain adaptation, which achieves competitive results.

📄 PDF Abstract BibTeX arXiv:1810.07455

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)ClassificationDialogue Act ClassificationDomain AdaptationGeneral Classificationspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Exploring the Viability of Synthetic Audio Data for Audio-Based Dialogue State Tracking

2023-12-04 · Jihyun Lee, Yejin Jeon, Wonjun Lee, Yunsu Kim 외

Dialogue state tracking plays a crucial role in extracting information in task-oriented dialogue systems. However, preceding research are limited to textual modalities, primarily due to the shortage of authentic human au…

Dialogue State TrackingTask-Oriented Dialogue Systems

Contextual Speech Extraction: Leveraging Textual History as an Implicit Cue for Target Speech Extraction

2025-03-11 · Minsu Kim, Rodrigo Mira, Honglie Chen, Stavros Petridis 외

In this paper, we investigate a novel approach for Target Speech Extraction (TSE), which relies solely on textual context to extract the target speech. We refer to this task as Contextual Speech Extraction (CSE). Unlike …

Speech Extraction

Exploring Speech Pattern Disorders in Autism using Machine Learning

2024-05-03 · Chuanbo Hu, Jacob Thrasher, Wenqi Li, Mindi Ruan 외

Diagnosing autism spectrum disorder (ASD) by identifying abnormal speech patterns from examiner-patient dialogues presents significant challenges due to the subtle and diverse manifestations of speech-related symptoms in…

DiagnosticregressionRhythm

InterroLang: Exploring NLP Models and Datasets through Dialogue-based Explanations

2023-10-09 · Nils Feldhus, Qianli Wang, Tatiana Anikina, Sahil Chopra 외

While recently developed NLP explainability methods let us open the black box in various ways (Madsen et al., 2022), a missing ingredient in this endeavor is an interactive tool offering a conversational interface. Such …

Dialogue Act ClassificationHate Speech DetectionQuestion Answering

SLIDE: Integrating Speech Language Model with LLM for Spontaneous Spoken Dialogue Generation

2025-01-01 · Haitian Lu, Gaofeng Cheng, Liuping Luo, Leying Zhang 외

Recently, ``textless" speech language models (SLMs) based on speech units have made huge progress in generating naturalistic speech, including non-verbal vocalizations. However, the generated speech samples often lack se…

Dialogue GenerationLanguage ModelingLanguage Modelling