Mining both Commonality and Specificity from Multiple Documents for Multi-Document Summarization
The multi-document summarization task requires the designed summarizer to generate a short text that covers the important information of original documents and satisfies content diversity. This paper proposes a multi-document summarization approach based on hierarchical clustering of documents. It utilizes the constructed class tree of documents to extract both the sentences reflecting the commonality of all documents and the sentences reflecting the specificity of some subclasses of these documents for generating a summary, so as to satisfy the coverage and diversity requirements of multi-document summarization. Comparative experiments with different variant approaches on DUC'2002-2004 datasets prove the effectiveness of mining both the commonality and specificity of documents for multi-document summarization. Experiments on DUC'2004 and Multi-News datasets show that our approach achieves competitive performance compared to the state-of-the-art unsupervised and supervised approaches.
Code (0)
등록된 구현이 없습니다.
Tasks
DiversityDocument SummarizationMulti-Document SummarizationSpecificitySimilar Papers 제목 키워드 기반
Comparative Document Analysis for Large Text Corpora
This paper presents a novel research problem on joint discovery of commonalities and differences between two individual documents (or document sets), called Comparative Document Analysis (CDA). Given any pair of document…
ArticlesStructure fusion based on graph convolutional networks for semi-supervised classification
Suffering from the multi-view data diversity and complexity for semi-supervised classification, most of existing graph convolutional networks focus on the networks architecture construction or the salient graph structure…
ClassificationGeneral ClassificationNode ClassificationSpecificityMulti-Source Uncertainty Mining for Deep Unsupervised Saliency Detection
Deep learning-based image salient object detection (SOD) heavily relies on large-scale training data with pixel-wise labeling. High-quality labels involve intensive labor and are expensive to acquire. In this paper, …
Deep Learningobject-detectionObject DetectionSaliency Detection+2Commonality and Individuality! Integrating Humor Commonality with Speaker Individuality for Humor Recognition
Humor recognition aims to identify whether a specific speaker's text is humorous. Current methods for humor recognition mainly suffer from two limitations: (1) they solely focus on one aspect of humor commonalities, igno…
Long Context is Not Long at All: A Prospector of Long-Dependency Data for Large Language Models
Long-context modeling capabilities are important for large language models (LLMs) in various applications. However, directly training LLMs with long context windows is insufficient to enhance this capability since some t…
AllComputational EfficiencySpecificity