PKUSUMSUM : A Java Platform for Multilingual Document Summarization
PKUSUMSUM is a Java platform for multilingual document summarization, and it sup-ports multiple languages, integrates 10 automatic summarization methods, and tackles three typical summarization tasks. The summarization platform has been released and users can easily use and update it. In this paper, we make a brief description of the char-acteristics, the summarization methods, and the evaluation results of the platform, and al-so compare PKUSUMSUM with other summarization toolkits.
Code (0)
등록된 구현이 없습니다.
Tasks
Chinese Word SegmentationDocument SummarizationInformation RetrievalSimilar Papers 제목 키워드 기반
Recommendations for Datasets for Source Code Summarization
Source Code Summarization is the task of writing short, natural language descriptions of source code. The main use for these descriptions is in software documentation e.g. the one-sentence Java method descriptions in Jav…
Code SummarizationSentenceSource Code SummarizationMultilingual Text Summarization on Financial Documents
This paper proposes a multilingual Automated Text Summarization (ATS) method targeting the Financial Narrative Summarization Task (FNS-2022). We developed two systems; the first uses a pre-trained abstractive summarizati…
Abstractive Text SummarizationText SummarizationMultiHumES: Multilingual Humanitarian Dataset for Extractive Summarization
When responding to a disaster, humanitarian experts must rapidly process large amounts of secondary data sources to derive situational awareness and guide decision-making. While these documents contain valuable informati…
Decision MakingExtractive SummarizationHumanitarianSupervising the Centroid Baseline for Extractive Multi-Document Summarization
The centroid method is a simple approach for extractive multi-document summarization and many improvements to its pipeline have been proposed. We further refine it by adding a beam search process to the sentence selectio…
Document SummarizationMulti-Document SummarizationSentenceEUR-Lex-Sum: A Multi- and Cross-lingual Dataset for Long-form Summarization in the Legal Domain
Existing summarization datasets come with two main drawbacks: (1) They tend to focus on overly exposed domains, such as news articles or wiki-like texts, and (2) are primarily monolingual, with few multilingual datasets.…
ArticlesForm