paper-with-me

Papers

HADES: Homologous Automated Document Exploration and Summarization

2023-02-25 · Piotr Wilczyński, Artur Żółkowski, Mateusz Krzyziński, Emilia Wiśnios, Bartosz Pieliński, Stanisław Giziński, Julian Sienkiewicz, Przemysław Biecek

This paper introduces HADES, a novel tool for automatic comparative documents with similar structures. HADES is designed to streamline the work of professionals dealing with large volumes of documents, such as policy documents, legal acts, and scientific papers. The tool employs a multi-step pipeline that begins with processing PDF documents using topic modeling, summarization, and analysis of the most important words for each topic. The process concludes with an interactive web app with visualizations that facilitate the comparison of the documents. HADES has the potential to significantly improve the productivity of professionals dealing with high volumes of documents, reducing the time and effort required to complete tasks related to comparative document analysis. Our package is publically available on GitHub.

📄 PDF Abstract BibTeX arXiv:2302.13099

Code (1)

mi2datalab/hades 공식 구현

Similar Papers 제목 키워드 기반

Re-evaluating Automatic Summarization with BLEU and 192 Shades of ROUGE

2015-09-01 · EMNLP 2015 9 · Yvette Graham
Machine Translation

Neural Natural Language Processing for Long Texts: A Survey on Classification and Summarization

2023-05-25 · Dimitrios Tsirmpas, Ioannis Gkionis, Georgios Th. Papadopoulos, Ioannis Mademlis

The adoption of Deep Neural Networks (DNNs) has greatly benefited Natural Language Processing (NLP) during the past decade. However, the demands of long document analysis are quite different from those of shorter texts, …

Document ClassificationDocument SummarizationManagementSentiment Analysis

Automated Metrics for Medical Multi-Document Summarization Disagree with Human Evaluations

2023-05-23 · Lucy Lu Wang, Yulia Otmakhova, Jay DeYoung, Thinh Hung Truong 외

Evaluating multi-document summarization (MDS) quality is difficult. This is especially true in the case of MDS for biomedical literature reviews, where models must synthesize contradicting evidence reported across differ…

Document SummarizationMulti-Document Summarization

Exploration of Summarization by Generative Language Models for Automated Scoring of Long Essays

2025-10-26 · Haowei Hua, Hong Jiao, Xinyi Wang arxiv

BERT and its variants are extensively explored for automated scoring. However, a limit of 512 tokens for these encoder-based models showed the deficiency in automated scoring of long essays. Thus, this research explores …

Automated Essay Scoring

An Empirical Survey on Long Document Summarization: Datasets, Models and Metrics

2022-07-03 · Huan Yee Koh, Jiaxin Ju, Ming Liu, Shirui Pan

Long documents such as academic articles and business reports have been the standard format to detail out important issues and complicated subjects that require extra attention. An automatic summarization system that can…

ArticlesDocument SummarizationText Summarization