paper-with-me

홈 › Papers

Progressive Training for Explainable Citation-Grounded Dialogue: Reducing Hallucination to Zero in English-Hindi LLMs

2026-03-19 · Vedant Pandya arxiv

Knowledge-grounded dialogue systems aim to generate informative, contextually relevant responses by conditioning on external knowledge sources. However, most existing approaches focus exclusively on English, lack explicit citation mechanisms for verifying factual claims, and offer limited transparency into model decision-making. We present XKD-Dial, a progressive four-stage training pipeline for explainable, knowledge-grounded dialogue generation in a bilingual (English-Hindi) setting, comprising: (1) multilingual adaptation, (2) English dialogue SFT with citation grounding, (3) bilingual dialogue SFT, and (4) GRPO alignment with citation-aware rewards. We evaluate six models spanning encoder-decoder (250M-3B) and decoder-only (1B-7B) architectures at every pipeline stage. Our key contributions are: (i) three post-hoc explainability analyses - cross-attention alignment, Integrated Gradients attribution, and occlusion-based causal grounding - applied systematically across the training trajectory to reveal how citation behaviour is learned, not only whether it is learned; (ii) citation-grounded SFT reduces hallucination to 0.0% for encoder-decoder models from Stage 2 onward; (iii) the progressive pipeline prevents catastrophic forgetting while improving Hindi capabilities; (iv) smaller models match larger models on English after SFT; and (v) GRPO provides marginal improvement over well-designed SFT for structured citation tasks. We evaluate across six automatic metrics (BLEU, ROUGE, BERTScore, FactScore, Citation-F1, and hallucination rate).

📄 PDF Abstract BibTeX arXiv:2603.18911

Code (0)

등록된 구현이 없습니다.

Tasks

Dialogue Generation

Similar Papers 제목 키워드 기반

A Grounded Interaction Protocol for Explainable Artificial Intelligence

2019-03-05 · Prashan Madumal, Tim Miller, Liz Sonenberg, Frank Vetere

Explainable Artificial Intelligence (XAI) systems need to include an explanation model to communicate the internal decisions, behaviours and actions to the interacting humans. Successful explanation involves both cogniti…

Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)

VDialogUE: A Unified Evaluation Benchmark for Visually-grounded Dialogue

2023-09-14 · Yunshui Li, Binyuan Hui, Zhaochao Yin, Wanwei He 외

Visually-grounded dialog systems, which integrate multiple modes of communication such as text and visual inputs, have become an increasingly popular area of investigation. However, the absence of a standardized evaluati…

Evaluating LLM-Generated Versus Human-Authored Responses in Role-Play Dialogues

2025-09-22 · Dongxu Lu, Johan Jeuring, Albert Gatt arxiv

Evaluating large language models (LLMs) in long-form, knowledge-grounded role-play dialogues remains challenging. This study compares LLM-generated and human-authored responses in multi-turn professional training simulat…

Facilitating Multi-turn Emotional Support Conversation with Positive Emotion Elicitation: A Reinforcement Learning Approach

2023-07-16 · Jinfeng Zhou, Zhuang Chen, Bo wang, Minlie Huang

Emotional support conversation (ESC) aims to provide emotional support (ES) to improve one's mental state. Existing works stay at fitting grounded responses and responding strategies (e.g., question), which ignore the ef…

Spatial AMR: Expanded Spatial Annotation in the Context of a Grounded Minecraft Corpus

2020-05-01 · LREC 2020 5 · Julia Bonn, Martha Palmer, Zheng Cai, Kristin Wright-Bettner

This paper presents an expansion to the Abstract Meaning Representation (AMR) annotation schema that captures fine-grained semantically and pragmatically derived spatial information in grounded corpora. We describe a new…

Abstract Meaning RepresentationMinecraftSentence