paper-with-me

Papers

DISCO: Document Intelligence Suite for COmparative Evaluation

2026-03-04 · Kenza Benkirane, Dan Goldwater, Martin Asenov, Aneiss Ghodsi arxiv

Document intelligence requires accurate text extraction and reliable reasoning over document content. We introduce \textbf{DISCO}, a \emph{Document Intelligence Suite for COmparative Evaluation}, that evaluates optical character recognition (OCR) pipelines and vision-language models (VLMs) separately on parsing and question answering across diverse document types, including handwritten text, multilingual scripts, medical forms, infographics, and multi-page documents. Our evaluation shows that performance varies substantially across tasks and document characteristics, underscoring the need for complexity-aware approach selection. OCR pipelines are generally more reliable for handwriting and for long or multi-page documents, where explicit text grounding supports text-heavy reasoning, while VLMs perform better on multilingual text and visually rich layouts. Task-aware prompting yields mixed effects, improving performance on some document types while degrading it on others. These findings provide empirical guidance for selecting document processing strategies based on document structure and reasoning demands.

📄 PDF Abstract BibTeX arXiv:2603.23511

Code (0)

등록된 구현이 없습니다.

Tasks

Question Answering

Similar Papers 제목 키워드 기반

A Test Suite for Evaluating Discourse Phenomena in Document-level Neural Machine Translation

2020-12-01 · AACL (iwdp) 2020 12 · Xinyi Cai, Deyi Xiong

The need to evaluate the ability of context-aware neural machine translation (NMT) models in dealing with specific discourse phenomena arises in document-level NMT. However, test sets that satisfy this need are rare. In …

Machine TranslationNMTTranslation

Discourse Cohesion Evaluation for Document-Level Neural Machine Translation

2022-08-19 · Xin Tan, Longyin Zhang, Guodong Zhou

It is well known that translations generated by an excellent document-level neural machine translation (NMT) model are consistent and coherent. However, existing sentence-level evaluation metrics like BLEU can hardly ref…

Machine TranslationNMTSentenceTranslation

A Test Suite and Manual Evaluation of Document-Level NMT at WMT19

2019-08-08 · Kateřina Rysová, Magdaléna Rysová, Tomáš Musil, Lucie Poláková 외

As the quality of machine translation rises and neural machine translation (NMT) is moving from sentence to document level translations, it is becoming increasingly difficult to evaluate the output of translation systems…

Machine TranslationNMTSentenceTranslation

A Test Suite and Manual Evaluation of Document-Level NMT at WMT19

2019-08-01 · WS 2019 8 · Kate{\v{r}}ina Rysov{\'a}, Magdal{\'e}na Rysov{\'a}, Tom{\'a}{\v{s}} Musil, Lucie Pol{\'a}kov{\'a} 외

As the quality of machine translation rises and neural machine translation (NMT) is moving from sentence to document level translations, it is becoming increasingly difficult to evaluate the output of translation systems…

Machine TranslationNMTSentenceTranslation

The Common Core Ontologies

2024-04-27 · Mark Jensen, Giacomo De Colle, Sean Kindya, Cameron More 외

The Common Core Ontologies (CCO) are designed as a mid-level ontology suite that extends the Basic Formal Ontology. CCO has since been increasingly adopted by a broad group of users and applications and is proposed as th…