paper-with-me

홈 › Papers

A Three-Stage Learning Framework for Low-Resource Knowledge-Grounded Dialogue Generation

2021-09-09 · EMNLP 2021 11 · Shilei Liu, Xiaofeng Zhao, Bochao Li, Feiliang Ren, Longhui Zhang, Shujuan Yin

Neural conversation models have shown great potentials towards generating fluent and informative responses by introducing external background knowledge. Nevertheless, it is laborious to construct such knowledge-grounded dialogues, and existing models usually perform poorly when transfer to new domains with limited training samples. Therefore, building a knowledge-grounded dialogue system under the low-resource setting is a still crucial issue. In this paper, we propose a novel three-stage learning framework based on weakly supervised learning which benefits from large scale ungrounded dialogues and unstructured knowledge base. To better cooperate with this framework, we devise a variant of Transformer with decoupled decoder which facilitates the disentangled learning of response generation and knowledge incorporation. Evaluation results on two benchmarks indicate that our approach can outperform other state-of-the-art methods with less training data, and even in zero-resource scenario, our approach still performs well.

📄 PDF Abstract BibTeX arXiv:2109.04096

Code (1)

neukg/kat-tslf 공식 구현 pytorch

Tasks

DecoderDialogue GenerationResponse GenerationWeakly-supervised Learning

Methods 이 논문이 사용한 방법론

Multi-Head Attention 설명 없음
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…
Position-Wise Feed-Forward Layer 설명 없음
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…

Similar Papers 제목 키워드 기반

ReFineG: Synergizing Small Supervised Models and LLMs for Low-Resource Grounded Multimodal NER

2025-09-13 · Jielong Tang, Shuang Wang, Zhenxing Wang, Jianxing Yu 외 arxiv

Grounded Multimodal Named Entity Recognition (GMNER) extends traditional NER by jointly detecting textual mentions and grounding them to visual regions. While existing supervised methods achieve strong performance, they …

Grounded Multimodal Named Entity RecognitionVisual Grounding

KnowPrefix-Tuning: A Two-Stage Prefix-Tuning Framework for Knowledge-Grounded Dialogue Generation

2023-06-27 · Jiaqi Bai, Zhao Yan, Jian Yang, Xinnian Liang 외

Existing knowledge-grounded conversation systems generate responses typically in a retrieve-then-generate manner. They require a large knowledge base and a strong knowledge retrieval component, which is time- and resourc…

Dialogue GenerationResponse GenerationRetrieval

OASIS: A Multilingual and Multimodal Dataset for Culturally Grounded Spoken Visual QA

2025-10-07 · Firoj Alam, Ali Ezzat Shahroor, Md. Arid Hasan, Zien Sheikh Ali 외 arxiv

Large-scale multimodal models achieve strong results on tasks like Visual Question Answering (VQA), but they are often limited when queries require cultural and visual information, everyday knowledge, particularly in low…

Visual Question AnsweringObject Recognition

A Unified Pre-training Framework for Conversational AI

2021-05-06 · Siqi Bao, Bingjin Chen, Huang He, Xin Tian 외

In this work, we explore the application of PLATO-2 on various dialogue systems, including open-domain conversation, knowledge grounded dialogue, and task-oriented conversation. PLATO-2 is initially designed as an open-d…

ChatbotInteractive Evaluation of DialogResponse Generation

Zero-Resource Knowledge-Grounded Dialogue Generation

2020-08-29 · NeurIPS 2020 12 · Linxiao Li, Can Xu, Wei Wu, Yufan Zhao 외

While neural conversation models have shown great potentials towards generating informative and engaging responses via introducing external knowledge, learning such a model often requires knowledge-grounded dialogues tha…

Dialogue Generation