Self-supervised Dialogue Learning for Spoken Conversational Question Answering
In spoken conversational question answering (SCQA), the answer to the corresponding question is generated by retrieving and then analyzing a fixed spoken document, including multi-part conversations. Most SCQA systems have considered only retrieving information from ordered utterances. However, the sequential order of dialogue is important to build a robust spoken conversational question answering system, and the changes of utterances order may severely result in low-quality and incoherent corpora. To this end, we introduce a self-supervised learning approach, including incoherence discrimination, insertion detection, and question prediction, to explicitly capture the coreference resolution and dialogue coherence among spoken documents. Specifically, we design a joint learning framework where the auxiliary self-supervised tasks can enable the pre-trained SCQA systems towards more coherent and meaningful spoken dialogue learning. We also utilize the proposed self-supervised learning tasks to capture intra-sentence coherence. Experimental results demonstrate that our proposed method provides more coherent, meaningful, and appropriate responses, yielding superior performance gains compared to the original pre-trained language models. Our method achieves state-of-the-art results on the Spoken-CoQA dataset.
Code (0)
등록된 구현이 없습니다.
Tasks
Conversational Question Answeringcoreference-resolutionCoreference ResolutionQuestion AnsweringSelf-Supervised LearningSentenceSimilar Papers 제목 키워드 기반
Towards Data Distillation for End-to-end Spoken Conversational Question Answering
In spoken question answering, QA systems are designed to answer questions from contiguous text spans within the related speech transcripts. However, the most natural way that human seek or test their knowledge is via hum…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Conversational Question AnsweringQuestion Answering+2Contextualized Attention-based Knowledge Transfer for Spoken Conversational Question Answering
Spoken conversational question answering (SCQA) requires machines to model complex dialogue flow given the speech utterances and text corpora. Different from traditional text question answering (QA) tasks, SCQA involves …
Audio Signal ProcessingConversational Question AnsweringKnowledge DistillationQuestion Answering+1Spoken Conversational Search for General Knowledge
We present a spoken conversational question answering proof of concept that is able to answer questions about general knowledge from Wikidata. The dialogue component does not only orchestrate various components but also …
Conversational Question AnsweringConversational SearchGeneral KnowledgeQuestion AnsweringEnd-to-end Spoken Conversational Question Answering: Task, Dataset and Model
In spoken question answering, the systems are designed to answer questions from contiguous text spans within the related speech transcripts. However, the most natural way that human seek or test their knowledge is via hu…
4kConversational Question AnsweringQuestion AnsweringSpoken Language Understanding+1Generative Spoken Dialogue Language Modeling
We introduce dGSLM, the first "textless" model able to generate audio samples of naturalistic spoken dialogues. It uses recent work on unsupervised spoken unit discovery coupled with a dual-tower transformer architecture…
Language ModelingLanguage Modelling