paper-with-me

홈 › Papers

Improved Evidence Extraction and Metrics for Document Inconsistency Detection with LLMs

2026-01-06 · Nelvin Tan, Yaowen Zhang, James Asikin Cheung, Fusheng Liu, Yu-Ching Shih, Dong Yang arxiv

Large language models (LLMs) are becoming useful in many domains due to their impressive abilities that arise from large training datasets and large model sizes. However, research on LLM-based approaches to document inconsistency detection is relatively limited. We address this gap by investigating evidence extraction capabilties of LLMs for document inconsistency detection. To this end, we introduce new comprehensive evidence-extraction metrics and a redact-and-retry framework with constrained filtering that substantially improves evidence extraction performance over other prompting methods. We support our approach with strong experimental results and release a new semi-synthetic dataset for evaluating evidence extraction.

📄 PDF Abstract BibTeX arXiv:2601.02627

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Entity and Evidence Guided Relation Extraction for DocRED

2020-08-27 · Kevin Huang, Guangtao Wang, Tengyu Ma, Jing Huang

Document-level relation extraction is a challenging task which requires reasoning over multiple sentences in order to predict relations in a document. In this paper, we pro-pose a joint training frameworkE2GRE(Entity and…

Document-level Relation ExtractionLanguage ModelingLanguage ModellingRelation+1

Revising FUNSD dataset for key-value detection in document images

2020-10-11 · Hieu M. Vu, Diep Thi-Ngoc Nguyen

FUNSD is one of the limited publicly available datasets for information extraction from document im-ages. The information in the FUNSD dataset is defined by text areas of four categories ("key", "value", "header", "other…

Making Document-Level Information Extraction Right for the Right Reasons

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Document-level information extraction is a flexible framework compatible with applications where information is not necessarily localized in a single sentence. For example, key features of a diagnosis in radiology a repo…

Sentence

From Chaos to Clarity: Schema-Constrained AI for Auditable Biomedical Evidence Extraction from Full-Text PDFs

2025-12-31 · Pouria Mortezaagha, Joseph Shaw, Bowen Sun, Arya Rahgozar arxiv

Biomedical evidence synthesis relies on accurate extraction of methodological, laboratory, and outcome variables from full-text research articles, yet these variables are embedded in complex scientific PDFs that make man…

Document AI

MDACE: MIMIC Documents Annotated with Code Evidence

2023-07-07 · ACL 2023 7 · Hua Cheng, Rana Jafari, April Russell, Russell Klopfer 외

We introduce a dataset for evidence/rationale extraction on an extreme multi-label classification task over long medical documents. One such task is Computer-Assisted Coding (CAC) which has improved significantly in rece…

Document ClassificationExtreme Multi-Label ClassificationMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATION