paper-with-me

DUC 2004

홈페이지 · 논문 15편

The DUC2004 dataset is a dataset for document summarization. Is designed and used for testing only. It consists of 500 news articles, each paired with four human written summaries. Specifically it consists of 50 clusters of Text REtrieval Conference (TREC) documents, from the following collections: AP newswire, 1998-2000; New York Times newswire, 1998-2000; Xinhua News Agency (English version), 1996-2000. Each cluster contained on average 10 documents. Source: Discrete Optimization for Unsupervised Sentence Summarization with Word-Level Extraction Image Source: https://duc.nist.gov/duc2004/

Texts English

벤치마크

Text Summarization on DUC 2004 Task 1 결과 13개
Extractive Text Summarization on DUC 2004 결과 2개
Extractive Text Summarization on DUC 2004 Task 1 결과 2개
Multi-Document Summarization on DUC 2004 결과 2개