paper-with-me

Papers Key-value Pair Extraction

“Key-value Pair Extraction” 태그가 달린 논문 13편 · 필터 해제

MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports

2026-05-04 · Yingyun Li, Yu Wang, Haiyang Qian arxiv

Semi-structured information extraction (IE) from OCR-derived clinical reports is crucial for efficiently reconstructing patients' longitudinal medical histories. In practice, this scenario commonly involves three tasks: …

Key-value Pair ExtractionInformation ExtractionQuestion Answering

KVP10k : A Comprehensive Dataset for Key-Value Pair Extraction in Business Documents

2024-05-01 · Oshri Naparstek, Roi Pony, Inbar Shapira, Foad Abo Dahood 외

In recent years, the challenge of extracting information from business documents has emerged as a critical task, finding applications across numerous domains. This effort has attracted substantial interest from both indu…

DiversityKey Information ExtractionKey-value Pair Extraction

UniVIE: A Unified Label Space Approach to Visual Information Extraction from Form-like Documents

2024-01-17 · Kai Hu, Jiawei Wang, WeiHong Lin, Zhuoyao Zhong 외

Existing methods for Visual Information Extraction (VIE) from form-like documents typically fragment the process into separate subtasks, such as key information extraction, key-value pair extraction, and choice group ext…

DecoderFormKey Information ExtractionKey-value Pair Extraction+2

PEneo: Unifying Line Extraction, Line Grouping, and Entity Linking for End-to-end Document Pair Extraction

2024-01-07 · Zening Lin, Jiapeng Wang, Teng Li, Wenhui Liao 외

Document pair extraction aims to identify key and value entities as well as their relationships from visually-rich documents. Most existing methods divide it into two separate tasks: semantic entity recognition (SER) and…

Key Information ExtractionKey-value Pair ExtractionRelation ExtractionSemantic entity labeling

Reading Order Matters: Information Extraction from Visually-rich Documents by Token Path Prediction

2023-10-17 · Chong Zhang, Ya Guo, Yi Tu, Huan Chen 외

Recent advances in multimodal pre-trained models have significantly improved information extraction from visually-rich documents (VrDs), in which named entity recognition (NER) is treated as a sequence-labeling task of p…

Entity LinkingKey Information ExtractionKey-value Pair Extractionnamed-entity-recognition+9

Towards Zero-shot Relation Extraction in Web Mining: A Multimodal Approach with Relative XML Path

2023-05-23 · Zilong Wang, Jingbo Shang

The rapid growth of web pages and the increasing complexity of their structure poses a challenge for web mining models. Web mining models are required to understand the semi-structured web pages, particularly when little…

Contrastive LearningKey-value Pair ExtractionRelationRelation Extraction

GeoLayoutLM: Geometric Pre-training for Visual Information Extraction

2023-04-21 · CVPR 2023 1 · Chuwei Luo, Changxu Cheng, Qi Zheng, Cong Yao

Visual information extraction (VIE) plays an important role in Document Intelligence. Generally, it is divided into two tasks: semantic entity recognition (SER) and relation extraction (RE). Recently, pre-trained models …

Document AIentity_extractionEntity LinkingKey Information Extraction+3

A Question-Answering Approach to Key Value Pair Extraction from Form-like Document Images

2023-04-17 · Kai Hu, Zhuoyuan Wu, Zhuoyao Zhong, WeiHong Lin 외

In this paper, we present a new question-answering (QA) based key-value pair extraction approach, called KVPFormer, to robustly extracting key-value relationships between entities from form-like document images. Specific…

DecoderFormKey-value Pair ExtractionPrediction+1

LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

2022-04-18 · Yupan Huang, Tengchao Lv, Lei Cui, Yutong Lu 외

Self-supervised pre-training techniques have achieved remarkable progress in Document AI. Most multimodal pre-trained models use a masked language modeling objective to learn bidirectional representations on the text mod…

cross-modal alignmentDocument AIdocument-image-classificationDocument Image Classification+16

LiLT: A Simple yet Effective Language-Independent Layout Transformer for Structured Document Understanding

2022-02-28 · ACL 2022 5 · Jiapeng Wang, Lianwen Jin, Kai Ding

Structured document understanding has attracted considerable attention and made significant progress recently, owing to its crucial role in intelligent document processing. However, most existing related models can only …

Document Image Classificationdocument understandingKey Information ExtractionKey-value Pair Extraction+1

OCR-free Document Understanding Transformer

2021-11-30 · Geewook Kim, Teakgyu Hong, Moonbin Yim, Jeongyeon Nam 외

Understanding document images (e.g., invoices) is a core but challenging task since it requires complex functions such as reading text and a holistic understanding of the document. Current Visual Document Understanding (…

Document Image Classificationdocument understandingKey-value Pair ExtractionOptical Character Recognition+2

LayoutXLM: Multimodal Pre-training for Multilingual Visually-rich Document Understanding

2021-04-18 · Yiheng Xu, Tengchao Lv, Lei Cui, Guoxin Wang 외

Multimodal pre-training with text, layout, and image has achieved SOTA performance for visually-rich document understanding tasks recently, which demonstrates the great potential for joint learning across different modal…

Document Image Classificationdocument understandingFormKey-value Pair Extraction

LayoutLMv2: Multi-modal Pre-training for Visually-Rich Document Understanding

2020-12-29 · ACL 2021 5 · Yang Xu, Yiheng Xu, Tengchao Lv, Lei Cui 외

Pre-training of text and layout has proved effective in a variety of visually-rich document understanding tasks due to its effective model architecture and the advantage of large-scale unlabeled scanned/digital-born docu…

Document Image ClassificationDocument Layout Analysisdocument understandingKey Information Extraction+7
1–13 / 13