paper-with-me

홈 › Papers

Learning to Select Context in a Hierarchical and Global Perspective for Open-domain Dialogue Generation

2021-02-18 · Lei Shen, Haolan Zhan, Xin Shen, Yang Feng

Open-domain multi-turn conversations mainly have three features, which are hierarchical semantic structure, redundant information, and long-term dependency. Grounded on these, selecting relevant context becomes a challenge step for multi-turn dialogue generation. However, existing methods cannot differentiate both useful words and utterances in long distances from a response. Besides, previous work just performs context selection based on a state in the decoder, which lacks a global guidance and could lead some focuses on irrelevant or unnecessary information. In this paper, we propose a novel model with hierarchical self-attention mechanism and distant supervision to not only detect relevant words and utterances in short and long distances, but also discern related information globally when decoding. Experimental results on two public datasets of both automatic and human evaluations show that our model significantly outperforms other baselines in terms of fluency, coherence, and informativeness.

📄 PDF Abstract BibTeX arXiv:2102.09282

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderDialogue GenerationInformativeness

Similar Papers 제목 키워드 기반

CHESS: Context-aware Hierarchical Efficient Semantic Selection for Long-Context LLM Inference

2026-02-24 · Chao Fei, Guozhong Li, Chenxi Liu, Panos Kalnis arxiv

Long-context LLMs demand accurate inference at low latency, yet decoding becomes primarily constrained by KV cache as context grows. Prior pruning methods are largely context-agnostic: their token selection ignores step-…

DGL-RSIS: Decoupling Global Spatial Context and Local Class Semantics for Training-Free Remote Sensing Image Segmentation

2025-08-30 · Boyi Li, Ce Zhang, Richard M. Timmerman, Wenxuan Bao arxiv

The emergence of vision language models (VLMs) bridges the gap between vision and language, enabling multimodal understanding beyond traditional visual-only deep learning models. However, transferring VLMs from the natur…

Referring Expression SegmentationSemantic SegmentationPrompt EngineeringImage Segmentation

Hierarchical Context-aware Network for Dense Video Event Captioning

2021-08-01 · ACL 2021 5 · Lei Ji, Xianglin Guo, Haoyang Huang, Xilin Chen

Dense video event captioning aims to generate a sequence of descriptive captions for each event in a long untrimmed video. Video-level context provides important information and facilities the model to generate consisten…

Descriptive

HiMu: Hierarchical Multimodal Frame Selection for Long Video Question Answering

2026-03-19 · Dan Ben-Ami, Gabriele Serussi, Kobi Cohen, Chaim Baskin arxiv

Long-form video question answering requires reasoning over extended temporal contexts, making frame selection a critical bottleneck for multi-modal large language models (MLLMs) bound by finite context windows. Within th…

Video Question AnsweringSpeech Recognition

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training

2025-07-30 · Xuecheng Wu, Danlei Huang, Heli Sun, Xinyi Yin 외 arxiv

Advances in Generative AI have made video-level deepfake detection increasingly challenging, exposing the limitations of current detection techniques. In this paper, we present HOLA, our solution to the Video-Level Deepf…

DeepFake Detection